Evaluating AI Agents and Autonomous Systems : Systematic Frameworks for Testing Autonomy, Tool-Calling Reliability, and Multi-Step Reasoning

Language: English

Published by Independently Published Mai 2026, 2026

9798196063763

Series: Book 2 of 5 - Architecting Enterprise Agents Series

  • Softcover
  • New
See all details

Seller: AHA-BUCH GmbH, Einbeck, GermanyAHA-BUCH GmbH

5-star seller

AbeBooks seller since August 14, 2006

View this seller's items
Softcover

Condition: New

£ 36.20

£ 26.15 shipping 
Ships from Germany to U.S.A.

Quantity: 2 available

Add to basket
Free 30-day returns

Item description from seller

Neuware - Evaluating AI Agents and Autonomous Systems: Systematic Frameworks for Testing Autonomy, Tool-Calling Reliability, and Multi-Step ReasoningAI agents are moving from impressive demos into real systems that call tools, retrieve data, make decisions, and execute workflows. But how do you know an autonomous agent is safe, reliable, and ready for production before it reaches users Evaluating AI Agents and Autonomous Systems gives engineers, architects, and technical leaders a practical framework for testing the systems that traditional software tests cannot fully capture. Built around autonomy, tool-calling reliability, multi-step reasoning, RAG evaluation, safety boundaries, observability, and multi-agent coordination, this book shows how to move from prompt testing to systematic agent validation. The book's structure covers evaluation harnesses, planning metrics, schema validation, LLM-as-a-judge workflows, RAG faithfulness, red teaming, trace analysis, human-in-the-loop review, scalable benchmarking, and MCP-based tool integration.Inside, readers will learn how to: - Measure whether an agent follows the right reasoning path, not just produces a polished answer.- Test tool selection, JSON/schema correctness, hallucinated tool calls, and recovery behavior.- Build evaluation pipelines for RAG, memory retrieval, multi-hop reasoning, and grounded tool arguments.- Apply red teaming, guardrails, PII audits, and boundary testing to autonomous workflows.- Use observability, tracing, regression tests, and human review to catch failures before deployment.For AI engineers, ML engineers, platform teams, and enterprise AI leaders, this book provides the testing discipline needed to ship agentic systems with confidence.

Seller Inventory # 9798196063763

Title
Evaluating AI Agents and Autonomous Systems : Systematic Frameworks for Testing Autonomy, Tool-Calling Reliability, and Multi-Step Reasoning
Author
Ethan Tyson
Publisher
Independently Published Mai 2026
Publication year
2026
Condition
Neu
Binding
Taschenbuch
Language
English
ISBN 13
9798196063763
Item weight
260 grams
Dimensions
254x178x8 mm
Series
Book 2 of 5: Architecting Enterprise Agents Series

AHA-BUCH GmbH

Einbeck, Germany

5-star seller

AbeBooks seller since August 14, 2006

Shipping rates from Germany to U.S.A.

Item5 to 7 business days7 to 10 business days
First item£ 26.15£ 26.15
Delivery times are set by sellers and vary by carrier and location. Orders passing through Customs may face delays and buyers are responsible for any associated duties or fees. Sellers may contact you regarding additional charges to cover any increased costs to ship your items.

Payment methods

  • Visa
  • Mastercard
  • American Express
  • Apple Pay
  • Google Pay
  • Bank Wire Transfer
  • Check
  • Paypal

Store description

Das Unternehmen AHA-BUCH GmbH: Seit der Gründung von AHA-BUCH im Juli 2005 ist unser Hauptziel, zufriedenen Kunden so schnell und so preisgünstig wie möglich ihren Bücherwunsch zu erfüllen. Unsere Firma beschäftigt 16 Mitarbeiter, die nur ein Ziel kennen: den Kunden und seine Wünsche! Auf über 3700 m2 Fläche haben wir über 100.000 Bücher, Modernes Antiquariat und Spiele auf Lager.

Specialty

Kinderbücher & Kinderhör Casetten, German Books, Software, Natur & Tiere, Ratgeber, Sachbücher, Englische Bücher, Medizin & Gesundheit, Universität & Studium

Seller's business information

AHA-BUCH GmbH

Garlebsen 48
Einbeck, Germany 37574