Everyone in the room saw the agent work. That is the problem.
In production it fails one run in twelve, never the same way twice, and you carry the pager. You are not stuck because the agent cannot do the work. You watched it do the work. You are stuck because doing it once, on stage, with you watching is the easy part, and doing it every time, while you sleep, is the job nobody named for you.
Harness Engineering names that gap: the demo cliff, and the climb back up. This is not a book about building an agent. You already built one. This is the book about making it deliver: the evals, verification, guardrails, observability, and recovery that turn an impressive toy into a production LLM application you would put your name on the pager for.
The engineer's toolkit for reliable AI agents that actually deliver:
Read it and you will measure your agent instead of demoing it, ship changes behind an eval gate, catch wrong answers before your users do, and answer the 3 AM page with a query instead of a guess. A reliable agent still fails. Its failures are rare, cheap, caught before the user, and survivable when they are not. That is software reliability engineering, the discipline SRE brought to servers, applied to agents. This book is the engineering.
For software, ML, and platform engineers who already run agentic loops and now own one in production. Part of the AI and Agentic Engineering series.
"Sinopsis" puede pertenecer a otra edición de este libro.
Librería: California Books, Miami, FL, Estados Unidos de America
Condición: New. Print on Demand. Nº de ref. del artículo: I-9798181738430
Cantidad disponible: Más de 20 disponibles
Librería: PBShop.store US, Wood Dale, IL, Estados Unidos de America
PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000. Nº de ref. del artículo: L2-9798181738430
Cantidad disponible: Más de 20 disponibles
Librería: Grand Eagle Retail, Bensenville, IL, Estados Unidos de America
Paperback. Condición: new. Paperback. Everyone in the room saw the agent work. That is the problem.In production it fails one run in twelve, never the same way twice, and you carry the pager. You are not stuck because the agent cannot do the work. You watched it do the work. You are stuck because doing it once, on stage, with you watching is the easy part, and doing it every time, while you sleep, is the job nobody named for you.Harness Engineering names that gap: the demo cliff, and the climb back up. This is not a book about building an agent. You already built one. This is the book about making it deliver: the evals, verification, guardrails, observability, and recovery that turn an impressive toy into a production LLM application you would put your name on the pager for.The engineer's toolkit for reliable AI agents that actually deliver: Reliability as a number: a success rate against a fixed task set, tracked run over run, so better is a measurement instead of a feeling.The eval gate: agent evals run like CI, blocking a bad change before it ships.The verification wall: cheap checks first, expensive ones behind them, catching wrong output before a user ever does.The recovery path for the failures you cannot prevent: retry with judgment, fall back, escalate, roll back.Run records: LLM observability and monitoring that capture what the agent actually did, so a bad run is a query instead of an archaeology dig.The reliability budget for one real agent: a target you can defend and the guardrails that hold it.Read it and you will measure your agent instead of demoing it, ship changes behind an eval gate, catch wrong answers before your users do, and answer the 3 AM page with a query instead of a guess. A reliable agent still fails. Its failures are rare, cheap, caught before the user, and survivable when they are not. That is software reliability engineering, the discipline SRE brought to servers, applied to agents. This book is the engineering.For software, ML, and platform engineers who already run agentic loops and now own one in production. Part of the AI and Agentic Engineering series. This item is printed on demand. Shipping may be from multiple locations in the US or from the UK, depending on stock availability. Nº de ref. del artículo: 9798181738430
Cantidad disponible: 1 disponibles
Librería: PBShop.store UK, Fairford, GLOS, Reino Unido
PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000. Nº de ref. del artículo: L2-9798181738430
Cantidad disponible: Más de 20 disponibles
Librería: CitiRetail, Stevenage, Reino Unido
Paperback. Condición: new. Paperback. Everyone in the room saw the agent work. That is the problem. In production it fails one run in twelve, and you carry the pager.You are not stuck because the agent cannot do the work. You watched it do the work. You are stuck because doing it once, on stage, with you watching is the easy part, and doing it every time, while you sleep, is the job nobody named for you. This book names it: the demo cliff. Harness Engineering is the climb back up.You already built an agent. This book is about making it deliver. Evals, verification, guardrails, observability, recovery: the scaffolding that turns an impressive toy into a system you would put your name on the pager for.A reliable agent still fails. Its failures are rare, cheap, caught before the user, and survivable when they are not. That is engineering, and this book is the engineering.What you buildReliability as a number: a success rate against a fixed task set, tracked run over run.An eval gate that blocks a bad change before it ships, run like CI.A verification wall that catches wrong output before a user does, cheap checks first, expensive ones behind them.A recovery path for the failures you cannot prevent: retry with judgment, fall back, escalate, roll back.Run records of what the agent actually did, so a bad run is a query instead of an archaeology dig.A reliability budget for one real agent: a target you can defend and the guardrails that hold it.Who it is forSoftware, ML, and platform engineers who already run agentic loops and now own one in production. You shipped the feature, the demo recording is still in the channel, and the 3 AM page lands on you. If that is your week, this is your book.About the authorWes Halloran writes practical books for working developers and engineering leaders in the shift to AI-agent development. He has shipped agents that failed in ways no demo hinted at, so this book trusts measured runs over impressive ones. He works from real tool behavior and verified workflows, and never accepts a green light he did not check.Part of the AI and Agentic Engineering series. This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability. Nº de ref. del artículo: 9798181738430
Cantidad disponible: 1 disponibles