Reliable Evaluations for LLMs and AI Agents : End-to-End Evaluation Frameworks for LLMs and Autonomous AI Agents

Langue : anglais

Edité par Springer Nature Switzerland AG, 2026

303226748X / 9783032267481

Vendeur : AHA-BUCH GmbH, Einbeck, AllemagneAHA-BUCH GmbH

Vendeur avec une évaluation de 5 étoiles

Vendeur AbeBooks depuis 14 août 2006

Livre broché

Etat: Neuf

EUR 67,30

EUR 35,00 expédition 
Expédition depuis Allemagne vers Etats-Unis

Quantité disponible : 2 disponibles

Ajouter au panier
Retours gratuits sous 30 jours

A propos de cet article

Druck auf Anfrage Neuware - Printed after ordering - This book gives practitioners a concrete, systematic framework for designing evals that make AI systems safe, robust, and customer-ready before they reach production. Drawing on real-world failures, from chatbots that went off the rails to shopping assistants that hallucinated product information, it shows how seemingly small evaluation gaps can cascade into legal, financial, and reputational crisis, and how to close those gaps with disciplined, systematic testing.Moving from foundational concepts to advanced practice, Reliable Evals for LLMs and AI Agents introduces the four core levers of effective evals: sets, templates, metrics, and evaluators. It then extends these to the unique challenges of autonomous AI agents, where systems perceive, reason, act, and adapt in iterative loops that demand fundamentally different eval approaches. Along the way, it guides readers through benchmark selection, custom eval set design, statistical rigor in metrics, human and LLM-as-a-judge rating strategies, and the infrastructure needed to automate evals at scale.For engineering leaders, applied researchers, data scientists, and product teams shipping LLM- and agent-powered experiences, this volume offers a blueprint for building eval flywheels that continuously improve AI quality. It shows how to progress from ad-hoc checks to production-grade eval systems, align model metrics with real user satisfaction, integrate offline evals with online A/B testing, and design accessible interfaces that democratize rigorous testing across an organization.…

N° de réf. du vendeur 9783032267481

Titre
Reliable Evaluations for LLMs and AI Agents : End-to-End Evaluation Frameworks for LLMs and Autonomous AI Agents
Auteur
Yueqing Wang
Éditeur
Springer Nature Switzerland AG
Année de publication
2026
État de l'article
Neu
Reliure
Taschenbuch
Langue
anglais
ISBN à 10 chiffres
303226748X
ISBN à 13 chiffres
9783032267481
Poids de l'article
318 grammes
Dimensions
235x155x12 mm

AHA-BUCH GmbH

Einbeck, Allemagne

Vendeur avec une évaluation de 5 étoiles

Vendeur AbeBooks depuis 14 août 2006

Frais d'expédition de Allemagne vers Etats-Unis

Article7 à 10 jours ouvrés5 à 7 jours ouvrés
Premier articleEUR 35,00EUR 45,00
Les délais de livraison sont fixés par les vendeurs et varient en fonction du transporteur et du lieu. Les commandes transitant par les douanes peuvent être retardées et les acheteurs sont responsables de tous les droits ou frais associés. Les vendeurs peuvent vous contacter au sujet de frais supplémentaires afin de couvrir toute augmentation des coûts d'expédition de vos articles.

Modes de paiement

  • Visa
  • Mastercard
  • American Express
  • Carte Bleue
  • Apple Pay
  • Google Pay
  • Chèque
  • Paypal
  • Virement bancaire

Description de la boutique

Das Unternehmen AHA-BUCH GmbH: Seit der Gründung von AHA-BUCH im Juli 2005 ist unser Hauptziel, zufriedenen Kunden so schnell und so preisgünstig wie möglich ihren Bücherwunsch zu erfüllen. Unsere Firma beschäftigt 16 Mitarbeiter, die nur ein Ziel kennen: den Kunden und seine Wünsche! Auf über 3700 m2 Fläche haben wir über 100.000 Bücher, Modernes Antiquariat und Spiele auf Lager.

Spécialité

Kinderbücher & Kinderhör Casetten, German Books, Software, Natur & Tiere, Ratgeber, Sachbücher, Englische Bücher, Medizin & Gesundheit, Universität & Studium

Profil professionnel du vendeur

AHA-BUCH GmbH

Garlebsen 48
Einbeck, Allemagne 37574