9798196063763 - evaluating ai agents and autonomous systems: systematic frameworks for testing autonomy, tool-calling reliability, and multi-step reasoning: 2 (architecting enterprise agents series) de tyson, ethan (5 resultados)
Idioma: Inglés
Editorial: Independently Published, 2026
- Tapa blanda
Librería: PBShop.store US, Wood Dale, IL, Estados Unidos de AmericaPBShop.store US
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 23,38
Gastos de envío gratisSe envía dentro de Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000.
Idioma: Inglés
Editorial: Independently Published, 2026
- Tapa blanda
Librería: PBShop.store UK, Fairford, GLOS, Reino UnidoPBShop.store UK
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 22,31
Envío por EUR 4,86Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
PAP. Condición: New. New Book. Shipped from UK. Established seller since 2000.
- Más imágenes
Idioma: Inglés
Editorial: Createspace Independent Publishing Platform Mai 2026, 2026
- Tapa blanda
Librería: AHA-BUCH GmbH, Einbeck, AlemaniaAHA-BUCH GmbH
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 27,57
Envío por EUR 61,30Se envía de Alemania a Estados Unidos de AmericaCantidad disponible: 2 disponibles
Taschenbuch. Condición: Neu. Neuware - Evaluating AI Agents and Autonomous Systems: Systematic Frameworks for Testing Autonomy, Tool-Calling Reliability, and Multi-Step ReasoningAI agents are moving from impressive demos into real systems that call tools, retrieve data, make decisions, and execute workflows. But how do you know…an autonomous agent is safe, reliable, and ready for production before it reaches users Evaluating AI Agents and Autonomous Systems gives engineers, architects, and technical leaders a practical framework for testing the systems that traditional software tests cannot fully capture. Built around autonomy, tool-calling reliability, multi-step reasoning, RAG evaluation, safety boundaries, observability, and multi-agent coordination, this book shows how to move from prompt testing to systematic agent validation. The book's structure covers evaluation harnesses, planning metrics, schema validation, LLM-as-a-judge workflows, RAG faithfulness, red teaming, trace analysis, human-in-the-loop review, scalable benchmarking, and MCP-based tool integration.Inside, readers will learn how to: - Measure whether an agent follows the right reasoning path, not just produces a polished answer.- Test tool selection, JSON/schema correctness, hallucinated tool calls, and recovery behavior.- Build evaluation pipelines for RAG, memory retrieval, multi-hop reasoning, and grounded tool arguments.- Apply red teaming, guardrails, PII audits, and boundary testing to autonomous workflows.- Use observability, tracing, regression tests, and human review to catch failures before deployment.For AI engineers, ML engineers, platform teams, and enterprise AI leaders, this book provides the testing discipline needed to ship agentic systems with confidence.
Idioma: Inglés
Editorial: Independently published, 2026
- Tapa blanda
- Impresión bajo demanda
Librería: California Books, Miami, FL, Estados Unidos de AmericaCalifornia Books
Contactar con el vendedorVendedor de 4 estrellasCondición: Nuevo
EUR 22,31
Gastos de envío gratisSe envía dentro de Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
Condición: New. Print on Demand.
Idioma: Inglés
Editorial: Independently Published, 2026
- Tapa blanda
- Impresión bajo demanda
Librería: CitiRetail, Stevenage, Reino UnidoCitiRetail
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 25,86
Envío por EUR 43,24Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: 1 disponibles
Paperback. Condición: new. Paperback. Evaluating AI Agents and Autonomous Systems: Systematic Frameworks for Testing Autonomy, Tool-Calling Reliability, and Multi-Step ReasoningAI agents are moving from impressive demos into real systems that call tools, retrieve data, make decisions, and execute workflows. But how do you know a…n autonomous agent is safe, reliable, and ready for production before it reaches users?Evaluating AI Agents and Autonomous Systems gives engineers, architects, and technical leaders a practical framework for testing the systems that traditional software tests cannot fully capture. Built around autonomy, tool-calling reliability, multi-step reasoning, RAG evaluation, safety boundaries, observability, and multi-agent coordination, this book shows how to move from prompt testing to systematic agent validation. The book's structure covers evaluation harnesses, planning metrics, schema validation, LLM-as-a-judge workflows, RAG faithfulness, red teaming, trace analysis, human-in-the-loop review, scalable benchmarking, and MCP-based tool integration.Inside, readers will learn how to: Measure whether an agent follows the right reasoning path, not just produces a polished answer.Test tool selection, JSON/schema correctness, hallucinated tool calls, and recovery behavior.Build evaluation pipelines for RAG, memory retrieval, multi-hop reasoning, and grounded tool arguments.Apply red teaming, guardrails, PII audits, and boundary testing to autonomous workflows.Use observability, tracing, regression tests, and human review to catch failures before deployment.For AI engineers, ML engineers, platform teams, and enterprise AI leaders, this book provides the testing discipline needed to ship agentic systems with confidence. This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.

