Local AI Engineering with Ollama (Paperback)

Idioma: inglés

Editorial: Faun.Dev, 2026

2488111082 / 9782488111089

  • Tapa blanda
  • Nuevo
Ver todos los detalles

Librería: Grand Eagle Retail, Bensenville, IL, Estados Unidos de AmericaGrand Eagle Retail

Vendedor de 5 estrellas

Vendedor de AbeBooks desde el 12 de octubre de 2005

Tapa blanda

Condición: Nuevo

EUR 44,51

 Gastos de envío gratis 
Se envía dentro de Estados Unidos de America

Cantidad disponible: 1 disponible

Añadir al carrito
Devoluciones gratuitas de 30 días

Descripción del artículo del vendedor

Paperback. The model you depend on lives on someone else's hardware. They can change the price, change the rules, or retire it entirely, and you cannot stop them.Local AI Engineering with Ollama is how you stop renting and start owning. You take the model, the price, and the rules back into your own hands: run any model you want, when you want, where you want, and change how it behaves without a meter running.This is a practical book for developers who can run a command and edit a file but have no Machine Learning degree and want none. It skips the marketing and jumps into building things that run, on hardware you already own, with the network unplugged. Every command was executed on a real machine, and every output you see (JSON responses, error messages, token counts, training logs) came from an actual session, not from documentation.This book moves in one direction: from running your first model to shipping an agent that runs on your own hardware. Each chapter ends with something working, and each skill below builds on the one before it. By the end you will be able to: Understand what a model is actually doing: Tokens, predictions, weights, embeddings, attention, and the KV cache, each tied to a setting you will change.Install Ollama and size your hardware honestly: Install the runtime and tell if a model fits your RAM or VRAM before downloading.Pick, pull, and manage models: Read the Ollama and Hugging Face GGUF repos, choose quantization, and manage disk and memory.Drive Ollama from its API: Run models over HTTP from your code, and read tokens-per-second to compare on numbers.Control the context window: Size it so the model stops forgetting, and see what gets sent each turn.Operate a model under real conditions: Tune temperature, top_p, top_k, penalties, seed, keep-alive, and concurrency.Package a custom model with a Modelfile: One job, the same way every time, shipped as a single artifact.Fine-tune a model on your own data: Train Granite for English-to-SQL with QLoRA and Unsloth, then export to GGUF.Build against the Python SDK: Build Python programs with typed responses, ending in a management CLI.Build a working chat loop and see why it forgets: Write a REPL, then watch it fail to recall the last turn.Give the conversation a memory: Resend a running message list so the assistant follows the conversation.Stream replies and accept multi-line input: Print tokens as they arrive, and take multi-line prompts.Keep long chats inside the context window: Drop the oldest turns so the prompt never overflows.Summarize old turns instead of dropping them: Condense earlier messages with a second model through LangChain.Cache replies in Redis: Return repeated questions instantly, cutting latency and wasted compute.Add long-term memory that survives restarts: Wire in mem0 to recall user facts across sessions.Give the model tools to fetch live data: Add function calling, guarded against inventing numbers.Source those tools from an external MCP server: Serve tools over MCP, turning M times N into M plus N.Put a graphical interface in front of Ollama: Run Open WebUI in Docker, chat with your documents, lock it down for a team.If you can run a command and edit a file, you are qualified! Downloadable code included.So what are you waiting for to stop renting, start owning, and get a model running Shipping may be from multiple locations in the US or from the UK, depending on stock availability.…

N° de ref. del artículo 9782488111089

Título
Local AI Engineering with Ollama (Paperback)
Autor
Aymen El Amri
Editorial
Faun.Dev
Año de publicación
2026
Estado
new
Encuadernación
Paperback
Idioma
inglés
ISBN 10
2488111082
ISBN 13
9782488111089

Grand Eagle Retail

Bensenville, IL, Estados Unidos de America

Vendedor de 5 estrellas

Vendedor de AbeBooks desde el 12 de octubre de 2005

Tarifas de envío en Estados Unidos de America

ArtículoDe 6 a 14 días hábilesDe 6 a 16 días hábiles
Primer artículoEUR 0,00EUR 0,00
Los plazos de entrega los establecen los vendedores y varían según el transportista y la ubicación. Los pedidos que pasan por la aduana pueden sufrir retrasos y los compradores son responsables de los aranceles o tarifas asociadas. Los vendedores pueden ponerse en contacto con usted en relación con cargos adicionales para cubrir cualquier aumento en los costes de envío de los artículos.

Métodos de pago

  • Visa
  • Mastercard
  • American Express
  • Carte Bleue
  • Apple Pay
  • Google Pay

Información empresarial del vendedor

APOLLO ONLINE CORP.

605 Geddes Street
Wilmington, DE Estados Unidos de America 19805