The Local AI Performance Handbook (Paperback)

Idioma: inglés

Editorial: Independently Published, 2026

9798195802172

Serie: Libro 4 de 5 - Architecting Enterprise Agents Series

  • Tapa blanda
  • Nuevo
Ver todos los detalles

Librería: Grand Eagle Retail, Bensenville, IL, Estados Unidos de AmericaGrand Eagle Retail

Vendedor de 5 estrellas

Vendedor de AbeBooks desde 12 de octubre de 2005

Ver los artículos de este vendedor
Tapa blanda

Condición: Nuevo

EUR 22,41

 Gastos de envío gratis 
Se envía dentro de Estados Unidos de America

Cantidad disponible: 1 disponibles

Añadir al carrito
Devoluciones gratuitas de 30 días

Descripción del artículo del vendedor

Paperback. The Local AI Performance Handbook: Optimizing Ollama for Multi-GPU and Hardware AccelerationLocal AI is powerful, but poor configuration can turn expensive hardware into a slow, unstable bottleneck. If your Ollama setup struggles with VRAM limits, weak token throughput, GPU underuse, long context slowdowns, or unreliable multi-user workloads, this handbook gives you the practical performance playbook you need.The Local AI Performance Handbook is a technical guide to building faster, more private, and more reliable Ollama systems across NVIDIA CUDA, AMD ROCm, Apple Silicon, WSL2, Docker, Kubernetes, and multi-GPU environments. It moves beyond basic local model setup and focuses on the engineering details that determine real-world performance: hardware acceleration, VRAM planning, quantization, request concurrency, private RAG, secure deployment, benchmarking, and production maintenance. The book's scope is reflected in its coverage of hardware-specific runtimes, memory engineering, multi-GPU scheduling, quantization, high-concurrency handling, private RAG, deployment, agentic workflows, and troubleshooting.Inside, readers will learn how to: Configure Ollama for CUDA, ROCm, Apple Silicon, Vulkan, Docker, and WSL2.Calculate model memory footprints and avoid out-of-memory failures.Tune VRAM usage, KV cache behavior, context windows, and quantization choices.Scale Ollama across multiple GPUs and isolate workloads with resource controls.Benchmark tokens per second, latency, GPU utilization, and system bottlenecks.Deploy private AI inference with Docker Compose, Kubernetes, health checks, and secure API access.Build faster private RAG and local agent workflows without depending on cloud APIs.For developers, AI engineers, homelab builders, and technical teams serious about private AI performance, this book turns Ollama from a simple local model runner into a tuned inference platform. This item is printed on demand. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.

N° de ref. del artículo 9798195802172

Título
The Local AI Performance Handbook (Paperback)
Autor
Ethan Tyson
Editorial
Independently Published
Año de publicación
2026
Estado
new
Encuadernación
Paperback
Idioma
inglés
ISBN 13
9798195802172
Serie
Libro 4 de 5: Architecting Enterprise Agents Series

Grand Eagle Retail

Bensenville, IL, Estados Unidos de America

Vendedor de 5 estrellas

Vendedor de AbeBooks desde 12 de octubre de 2005

Tarifas de envío en Estados Unidos de America

ArtículoDe 6 a 14 días hábilesDe 6 a 16 días hábiles
Primer artículoEUR 0,00EUR 0,00
Los plazos de entrega los establecen los vendedores y varían según el transportista y la ubicación. Los pedidos que pasan por la aduana pueden sufrir retrasos y los compradores son responsables de los aranceles o tarifas asociadas. Los vendedores pueden ponerse en contacto con usted en relación con cargos adicionales para cubrir cualquier aumento en los costes de envío de los artículos.

Métodos de pago

  • Visa
  • Mastercard
  • American Express
  • Carte Bleue
  • Apple Pay
  • Google Pay

Información empresarial del vendedor

APOLLO ONLINE CORP.

605 Geddes Street
Wilmington, DE Estados Unidos de America 19805