Articles liés à Production Agentic AI Engineering Handbook: Production...

Production Agentic AI Engineering Handbook: Production Architecture, Delivery, Reliability, Security, Observability, Incident Operations, Performance, ... and FinOps (Agentic AI Engineering Series) - Couverture souple

Livre 10 sur 10: Agentic AI Engineering Series

Yusuf, Funke R.

 
9798193473374: Production Agentic AI Engineering Handbook: Production Architecture, Delivery, Reliability, Security, Observability, Incident Operations, Performance, ... and FinOps (Agentic AI Engineering Series)

Synopsis

Take AI agents from deployment to dependable production operations.

Production Agentic AI Engineering Handbook is Book 10 of the Agentic AI Engineer Series. It provides a practical engineering framework for deploying, operating, securing, scaling, monitoring, and governing agentic AI systems in real production environments.

Moving an agent into production requires much more than hosting a model endpoint. Production systems must manage configuration, secrets, deployment pipelines, state, queues, dependencies, failures, incidents, cost, security, privacy, observability, recovery, capacity, and continuous change.

This book shows how to engineer those concerns as part of a complete production operating model.

You will learn how to distinguish prototypes, products, platforms, and production systems; classify workloads by users, data, criticality, and risk; define service objectives and error budgets; and make evidence-based go/no-go decisions through structured production-readiness reviews.

The book then develops the delivery and platform foundations required for reliable operations, including runtime topologies, containers, infrastructure as code, supply-chain integrity, feature flags, secret management, environment promotion, and progressive delivery using blue-green, canary, shadow, and staged releases.

What You Will Learn
  • Design production reference architectures for agentic systems.
  • Define service objectives, error budgets, risk appetite, and responsibility boundaries.
  • Build repeatable deployment pipelines and controlled environment promotion.
  • Manage model, prompt, tool, schema, configuration, and data migrations.
  • Apply blue-green, canary, shadow, and progressive delivery strategies.
  • Limit blast radius using bulkheads, circuit breakers, retry budgets, and graceful degradation.
  • Manage queues, backpressure, admission control, and load shedding.
  • Design state durability, backup, restore, reconciliation, high availability, and disaster recovery.
  • Perform capacity testing, chaos exercises, and recovery drills.
  • Implement workload identity, least privilege, sandboxing, egress controls, and tenant isolation.
  • Apply encryption, data minimization, retention, deletion, auditability, and privacy controls.
  • Build telemetry across models, tools, workflows, infrastructure, and human interactions.
  • Create dashboards, alerts, SLOs, runbooks, incident triage, containment, and safe shutdown procedures.
  • Engineer end-to-end latency, caching, batching, concurrency, routing, retrieval, and tool-call efficiency.
  • Control token, storage, network, tool, and human operating costs.
  • Establish unit economics, budgets, quotas, chargeback, and FinOps guardrails.
  • Define SRE, platform, product, support, security, and risk responsibilities.
Build as You Learn

Build 1: A Production Readiness Dossier
Build 2: A Repeatable Agent Deployment Pipeline
Build 3: A Reliability and Recovery Plan
Build 4: A Security and Privacy Control Matrix
Build 5: An Agent Operations and Incident Console
Build 6: A Performance and FinOps Control Loop

The capstone brings everything together in a Production Agent Platform and Operating Model that combines golden paths, shared services, release evidence, evaluation gates, reliability controls, security, privacy, recovery, observability, cost management, incident operations, and governance.

Written for AI engineers, software engineers, cloud and platform architects, SRE teams, security professionals, technical leaders, and organizations preparing to operate AI agents at scale, this handbook provides the practical bridge between a working agent and a dependable production service.

Les informations fournies dans la section « Synopsis » peuvent faire référence à une autre édition de ce titre.