Workload optimization gpus cuda de maranto steven (4 resultados)

- Tapa blanda
Librería: PBShop.store US, Wood Dale, IL, Estados Unidos de AmericaPBShop.store US
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 25,65
Gastos de envío gratisSe envía dentro de Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000.

- Tapa blanda
Librería: PBShop.store UK, Fairford, GLOS, Reino UnidoPBShop.store UK
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 23,25
Envío por EUR 4,85Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000.

- Tapa blanda
- Impresión bajo demanda
Librería: California Books, Miami, FL, Estados Unidos de AmericaCalifornia Books
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 23,10
Gastos de envío gratisSe envía dentro de Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
Condición: New. Print on Demand.

- Tapa blanda
- Impresión bajo demanda
Librería: CitiRetail, Stevenage, Reino UnidoCitiRetail
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 27,00
Envío por EUR 43,14Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: 1 disponibles
Paperback. Condición: new. Paperback. AI Workload Optimization with GPUs, CUDA, and PyTorch: A Practical Guide to Faster Training, Lower Inference Latency, Better Throughput, and Scalable DeploymentYour AI model may work-but is it fast enough, efficient enough, and stable enough to survive real training and production use?Slow t…raining jobs, idle GPUs, memory crashes, weak throughput, high inference latency, and expensive cloud runs can turn a promising AI project into a costly engineering problem. Adding more hardware is not always the answer. If you do not know where the bottleneck is, you may waste time tuning the wrong part of the system.AI Workload Optimization with GPUs, CUDA, and PyTorch gives you a practical, measurement-first workflow for improving AI performance without guesswork. Built around the baseline, profile, optimize, verify method, this book helps you identify what is slowing down your workload, apply the right optimization, and confirm the result with clear metrics.Inside, you will learn how to: Benchmark training and inference correctlyProfile PyTorch workloads before changing codeImprove GPU utilization, memory use, and data loadingApply mixed precision, torch.compile, and CUDA-aware optimization carefullyScale training across multiple GPUsOptimize inference with PyTorch, ONNX Runtime, TensorRT, Triton, and vLLMMeasure latency, throughput, tail latency, tokens per second, and costThis book is written for machine learning engineers, software engineers, data scientists, AI infrastructure builders, and students who want practical GPU performance skills. The examples are self-contained, with code, commands, scripts, and project materials included directly in the book-no external companion repository required. This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.