Local AI Engineering with Ollama (Paperback)

Idioma: inglés

Editorial: Faun.Dev, 2026

2488111082 / 9782488111089

  • Tapa blanda
  • Nuevo
Ver todos los detalles

Librería: CitiRetail, Stevenage, Reino UnidoCitiRetail

Vendedor de 5 estrellas

Vendedor de AbeBooks desde el 29 de junio de 2022

Tapa blanda

Condición: Nuevo

EUR 43,84

Envío por EUR 43,16 
Se envía de Reino Unido a Estados Unidos de America

Cantidad disponible: 1 disponible

Añadir al carrito
Devoluciones gratuitas de 30 días

Descripción del artículo del vendedor

Paperback. The model you depend on lives on someone else's hardware. They can change the price, change the rules, or retire it entirely, and you cannot stop them.Local AI Engineering with Ollama is how you stop renting and start owning. You take the model, the price, and the rules back into your own hands: run any model you want, when you want, where you want, and change how it behaves without a meter running.This is a practical book for developers who can run a command and edit a file but have no Machine Learning degree and want none. It skips the marketing and jumps into building things that run, on hardware you already own, with the network unplugged. Every command was executed on a real machine, and every output you see (JSON responses, error messages, token counts, training logs) came from an actual session, not from documentation.This book moves in one direction: from running your first model to shipping an agent that runs on your own hardware. Each chapter ends with something working, and each skill below builds on the one before it. By the end you will be able to: Understand what a model is actually doing: Tokens, predictions, weights, embeddings, attention, and the KV cache, each tied to a setting you will change.Install Ollama and size your hardware honestly: Install the runtime and tell if a model fits your RAM or VRAM before downloading.Pick, pull, and manage models: Read the Ollama and Hugging Face GGUF repos, choose quantization, and manage disk and memory.Drive Ollama from its API: Run models over HTTP from your code, and read tokens-per-second to compare on numbers.Control the context window: Size it so the model stops forgetting, and see what gets sent each turn.Operate a model under real conditions: Tune temperature, top_p, top_k, penalties, seed, keep-alive, and concurrency.Package a custom model with a Modelfile: One job, the same way every time, shipped as a single artifact.Fine-tune a model on your own data: Train Granite for English-to-SQL with QLoRA and Unsloth, then export to GGUF.Build against the Python SDK: Build Python programs with typed responses, ending in a management CLI.Build a working chat loop and see why it forgets: Write a REPL, then watch it fail to recall the last turn.Give the conversation a memory: Resend a running message list so the assistant follows the conversation.Stream replies and accept multi-line input: Print tokens as they arrive, and take multi-line prompts.Keep long chats inside the context window: Drop the oldest turns so the prompt never overflows.Summarize old turns instead of dropping them: Condense earlier messages with a second model through LangChain.Cache replies in Redis: Return repeated questions instantly, cutting latency and wasted compute.Add long-term memory that survives restarts: Wire in mem0 to recall user facts across sessions.Give the model tools to fetch live data: Add function calling, guarded against inventing numbers.Source those tools from an external MCP server: Serve tools over MCP, turning M times N into M plus N.Put a graphical interface in front of Ollama: Run Open WebUI in Docker, chat with your documents, lock it down for a team.If you can run a command and edit a file, you are qualified! Downloadable code included.So what are you waiting for to stop renting, start owning, and get a Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.…

N° de ref. del artículo 9782488111089

Título
Local AI Engineering with Ollama (Paperback)
Autor
Aymen El Amri
Editorial
Faun.Dev
Año de publicación
2026
Estado
new
Encuadernación
Paperback
Idioma
inglés
ISBN 10
2488111082
ISBN 13
9782488111089

CitiRetail

Stevenage, Reino Unido

Vendedor de 5 estrellas

Vendedor de AbeBooks desde el 29 de junio de 2022

Tarifas de envío de Reino Unido a Estados Unidos de America

ArtículoDe 7 a 14 días hábilesDe 7 a 60 días hábiles
Primer artículoEUR 43,16EUR 43,16
Los plazos de entrega los establecen los vendedores y varían según el transportista y la ubicación. Los pedidos que pasan por la aduana pueden sufrir retrasos y los compradores son responsables de los aranceles o tarifas asociadas. Los vendedores pueden ponerse en contacto con usted en relación con cargos adicionales para cubrir cualquier aumento en los costes de envío de los artículos.

Métodos de pago

  • Visa
  • Mastercard
  • American Express
  • Carte Bleue
  • Apple Pay
  • Google Pay

Descripción de la tienda

Online business

Información empresarial del vendedor

ABC BOOKS LIMITED

10 John Street
London, Reino Unido WC1N 2EB