Napačna izbira? Nič za to! Izdelke lahko vrnete do 30 dni
Z darilnim bonom ne morete zgrešiti. Obdarovanec lahko v zameno za darilni bon izbere karkoli iz naše ponudbe.
Do 30 dni za vračilo
Evaluating AI Agents and Autonomous Systems:Systematic Frameworks for Testing Autonomy, Tool-Calling Reliability, and Multi-Step Reasoning
AI agents are moving from impressive demos into real systems that call tools, retrieve data, make decisions, and execute workflows. But how do you know an autonomous agent is safe, reliable, and ready for production before it reaches users?
Evaluating AI Agents and Autonomous Systems gives engineers, architects, and technical leaders a practical framework for testing the systems that traditional software tests cannot fully capture. Built around autonomy, tool-calling reliability, multi-step reasoning, RAG evaluation, safety boundaries, observability, and multi-agent coordination, this book shows how to move from prompt testing to systematic agent validation. The book's structure covers evaluation harnesses, planning metrics, schema validation, LLM-as-a-judge workflows, RAG faithfulness, red teaming, trace analysis, human-in-the-loop review, scalable benchmarking, and MCP-based tool integration.
Inside, readers will learn how to:
For AI engineers, ML engineers, platform teams, and enterprise AI leaders, this book provides the testing discipline needed to ship agentic systems with confidence.
Pozdravljeni! Sem Libroamiko, vaš knjižni svetovalec.
Kako vam lahko pomagam?