Download the free Kindle app and start reading Kindle books instantly on your smartphone, tablet, or computer - no Kindle device required.
Read instantly on your browser with Kindle for Web.
Using your mobile phone camera - scan the code below and download the Kindle app.
Follow the author
OK
AI Agents: The Definitive Guide: Design, Deployment, and Evaluation for Production
Purchase options and add-ons
As AI agents move from research labs into production, engineering teams face mounting challenges—fragile tools, rising inference costs, erratic behavior, and unmet expectations. AI Agents: The Definitive Guide addresses what most books avoid: how to design, deploy, and maintain real AI agents that actually work in the real world, not just in demos. Written by Nicole Koenigstein, this book offers the practical, system-level foundations needed to build robust, scalable, and secure agentic systems.
Whether you're tasked with making agent prototypes production-ready or building mission-critical automation from the ground up, this book guides you through every layer of the stack, without being framework-dependent. It covers architectures, tool integration, performance optimization, safety strategies, and advanced evaluation, with a relentless focus on reliability and long-term value.
- Design stateful, reasoning, and multi-agent systems
- Apply reinforcement learning, search, and test-time compute
- Build reliable tool integration and execution boundaries
- Evaluate agents across development and production
- Design memory, monitoring, fallbacks, and efficient infrastructure
- Secure agents through isolation, governance, and threat modeling
- ISBN-13979-8341666931
- Edition1st
- PublisherO'Reilly Media
- Publication dateOctober 6, 2026
- LanguageEnglish
- Dimensions7 x 0.78 x 9.19 inches
- Print length376 pages
Frequently bought together

Customers who viewed this item also viewed
- An Illustrated Guide to AI Agents: Concepts and Code for Building Agents with LLMs, Tools, and MemoryPaperbackFREE Shipping by AmazonThis title will be released on October 6, 2026.
- AI Engineering: Building Applications with Foundation ModelsPaperbackFREE Shipping by AmazonGet it as soon as Sunday, Sep 20
- Building Applications with AI Agents: Designing and Implementing Multiagent SystemsPaperbackFREE Shipping by AmazonGet it as soon as Sunday, Sep 20
- AI Agents in Action, Second Edition: Intelligent workflows with LLMs, MCP, A2A, and morePaperbackFREE Shipping by AmazonGet it as soon as Monday, Sep 21
- The Agentic AI Bible: The Complete and Up-to-Date Guide to Design, Develop, and Scale Goal-Driven, LLM-Powered Agents that Think, Execute and EvolvePaperbackFREE Shipping on orders over $35 shipped by AmazonGet it as soon as Monday, Sep 21
- Generative AI Design Patterns: Solutions to Common Challenges When Building GenAI Agents and ApplicationsPaperbackFREE Shipping by AmazonGet it as soon as Sunday, Sep 20
Customers also bought or read
- An Illustrated Guide to AI Agents: Concepts and Code for Building Agents with LLMs, Tools, and Memory
PreorderPaperback$79.99$79.99 - Building Applications with AI Agents: Designing and Implementing Multiagent Systems
Paperback$56.44$56.44FREE delivery Sun, Sep 20 - Agentic Coding with Claude Code: The everyday developer's guide to agentic coding with Claude Code
Paperback$36.33$36.33FREE delivery Mon, Sep 21 - Design Multi-Agent AI Systems Using MCP and A2A: Engineer your own Python-based agentic AI framework with tool use, memory, and multi-agent workflows
Paperback$41.79$41.79FREE delivery Mon, Sep 21 - AI Engineering: Building Applications with Foundation Models#1 Best SellerEnterprise Applications
Paperback$52.40$52.40FREE delivery Sun, Sep 20 - Vision Language Models: Building VLMs with Hugging Face
Just releasedPaperback$70.87$70.87FREE delivery Sun, Sep 20 - Building AI Agents with LLMs, RAG, and Knowledge Graphs: A practical guide to autonomous and modern AI agents
Paperback$41.99$41.99FREE delivery Mon, Sep 21 - AI Systems Performance Engineering: Optimizing Model Training and Inference Workloads with GPUs, CUDA, and PyTorch#1 Best SellerComputer Hardware Design & Architecture
Paperback$86.07$86.07FREE delivery Mon, Sep 21 - 30 Agents Every AI Engineer Must Build: Build production-ready agent systems using proven architectures and patterns
Paperback$41.79$41.79FREE delivery Mon, Sep 21 - AI Agents and Applications: With LangChain, LangGraph, and MCP
Paperback$51.69$51.69FREE delivery Mon, Sep 21 - Designing Machine Learning Systems: An Iterative Process for Production-Ready Applications
Paperback$40.00$40.00FREE delivery Sun, Sep 20 - Optimization: A Bootcamp for Machine Learning, Inverse Problems, and Control
Just releasedHardcover$63.86$63.86FREE delivery Sun, Sep 20 - Natural Language Processing with Transformers, Revised Edition
Paperback$41.60$41.60FREE delivery Sun, Sep 20 - The Infinity Machine: Demis Hassabis, DeepMind, and the Quest for Superintelligence#1 Best SellerKnowledge Capital
Hardcover$21.45$21.45Delivery Sun, Sep 20 - Agentic Design Patterns: A Hands-On Guide to Building Intelligent Systems
Paperback$35.53$35.53FREE delivery Sun, Sep 20 - Building AI Agents for Finance: Build and deploy robust financial agentic systems with advanced reasoning, architectures, and Python#1 New ReleaseFinancial Engineering
Paperback$41.99$41.99FREE delivery Mon, Sep 21 - Hands-On Machine Learning with Scikit-Learn and PyTorch: Concepts, Tools, and Techniques to Build Intelligent Systems
Paperback$77.72$77.72FREE delivery Mon, Sep 21 - Generative AI Design Patterns: Solutions to Common Challenges When Building GenAI Agents and Applications
Paperback$68.99$68.99FREE delivery Sun, Sep 20 - Context Engineering for Multi-Agent Systems: Move beyond prompting to build a Context Engine, a transparent architecture of context and reasoning
Paperback$37.99$37.99FREE delivery Mon, Sep 21 - Fine-Tuning AI: Customizing Large Language Models
PreorderPaperback$79.99$79.99FREE delivery Fri, Nov 27 - Designing Multi-Agent Systems: Principles, Patterns, and Implementation for AI Agents
Paperback$46.39$46.39FREE delivery Mon, Sep 21 - Build a Reasoning Model (From Scratch)#1 Best SellerProgramming Algorithms
Paperback$47.82$47.82FREE delivery Sun, Sep 20 - Deep Learning with PyTorch, Second Edition: Training and applying deep learning and generative AI models
Paperback$51.68$51.68FREE delivery Mon, Sep 21 - Designing AI Interfaces: Design Principles for Creative and Autonomous AI
Paperback$70.87$70.87FREE delivery Sun, Sep 20 - Agentic Architectural Patterns for Building Multi-Agent Systems: Proven design patterns and practices for GenAI, agents, RAG, LLMOps, and enterprise-scale AI systems
Paperback$44.99$44.99FREE delivery Mon, Sep 21 - Human Compatible: Artificial Intelligence and the Problem of Control
Paperback$13.61$13.61Delivery Mon, Sep 21 - Generative AI with LangChain: Build production-ready LLM applications and advanced agents using Python, LangChain, and LangGraph
Paperback$41.99$41.99FREE delivery Mon, Sep 21
From the brand
-
Machine Learning, AI & more
-
Machine Learning
-
Artificial Intelligence
-
Deep Learning
-
Language Processing (NLP, LLM)
-
Sharing the knowledge of experts
O'Reilly's mission is to change the world by sharing the knowledge of innovators. For over 40 years, we've inspired companies and individuals to do new things (and do them better) by providing the skills and understanding that are necessary for success.
Our customers are hungry to build the innovations that propel the world forward. And we help them do just that.
From the Publisher
From the Preface
What This Book Is About
This book is a practical systems guide to designing, building, deploying, and evaluating AI agents for production.
It begins with the architectural foundations that make agentic behavior possible: state, control flow, planning, reasoning, reflection, and action. From there, it moves into increasingly complex systems, including human-in-the-loop workflows, hierarchical agents, swarms, search-based reasoning, reinforcement learning, and test-time compute.
The book also looks at the models behind the agents and the trade-offs involved in choosing between decoder-only, encoder-only, mixture-of-experts, reasoning, open-weight, and closed models. You’ll see how model choice, quantization, adapters, and inference constraints shape the behavior and feasibility of the larger system.
A substantial part of the book focuses on the transition from prototype to production. You’ll learn how to structure data contracts, validate outputs, integrate tools, use the Model Context Protocol, isolate execution, and govern programmatic tool calls. The later chapters address deployment, monitoring, fallbacks, inference backends, caching, evaluation, memory, infrastructure, cost, and threat modeling.
Evaluation is treated as part of the full agent lifecycle. This includes stress testing before deployment, observing behavior in production, turning traces into custom benchmarks, and assessing long-horizon reasoning, tool use, and system reliability.
Throughout the book, theory is paired with concrete implementations, architectural patterns, production considerations, and the trade-offs that become visible only when agentic systems are expected to operate under real conditions.
What This Book Is Not
This is not an introduction to large language models (LLMs). I assume that you already understand the basics of LLMs, embeddings, inference, and modern language models. Relevant concepts are explained where necessary, but the focus is on how those models become part of larger agentic systems.
This is also not a prompt engineering guide or a collection of API recipes. Prompts and instructions matter, but they’re only one part of an agent. This book provides the foundation for designing, evaluating, securing, and deploying production agents, including state, planning, tools, contracts, memory, execution boundaries, infrastructure, and governance.
Nor is this a framework manual. LangGraph, Pydantic, MCP, inference backends, and other tools appear throughout the book to make the examples concrete, but my goal is not to teach you one implementation stack. Frameworks and interfaces will change. My emphasis is on the architectural and engineering principles that allow you to evaluate new tools and apply them to your own systems.
Finally, this book is also not a collection of polished demonstrations that stop once an agent successfully calls a tool. It focuses on what happens or fails after your prototype works: reliability, observability, failure recovery, security, cost, deployment, and long-term maintainability.
Who This Book Is For
This book is written for intermediate to advanced software engineers, data scientists, machine learning engineers, MLOps engineers, AI and DevOps specialists, and AI infrastructure engineers who are responsible for designing, building, deploying, and maintaining LLM-powered agents in production. It’s for readers dealing with the realities that appear once an agent moves beyond a controlled prototype.
Specifically, I wrote this for:
- The engineer who needs to turn an agent prototype into a reliable production system and troubleshoot what happens when state breaks, tools fail, latency increases, or costs begin to scale
- The AI infrastructure engineers who are directly responsible for designing, developing, deploying, and maintaining AI agents powered by LLMs in production environments
- The architect who needs to decide how models, tools, memory, control flow, evaluation, security, and infrastructure should work together across the full system
- The technical leader who needs to assess feasibility, estimate cost, guide implementation teams, and understand the trade-offs between performance, reliability, security, and autonomy
- The deep-tech founder or CTO who needs enough technical depth to evaluate solutions, set realistic expectations, and make decisions about how an agentic product can be built and operated sustainably
A working understanding of Python and large language models is assumed. Experience with cloud platforms can be helpful, but isn’t required. It’s also useful if you understand how transformers process sequences and have at least a basic sense of how embeddings work or vector stores support retrieval.
Throughout these chapters, I don’t shy away from the complexity that production agent systems demand. Technical depth, however, will always serve a practical purpose and I pair it with clear implementations and concrete examples. The book focuses on the details that emerge when agents are expected to operate under real conditions, because those are the details that determine whether a system merely works in a demo or can be maintained in production.
Editorial Reviews
About the Author
She served as an external evaluator for a European Commission AI Grand Challenge and has advised IOSCO on generative AI in regulated environments. She also serves on advisory boards for leading AI and quantitative finance conferences. Nicole regularly delivers invited talks and technical workshops across academia, industry, and international events.
She is the author of Math for Machine Learning and Transformers in Action with Manning Publications. Her forthcoming books, Transformers: The Definitive Guide--Applications Beyond NLP and AI Agents: The Definitive Guide, will be published by O'Reilly Media.
Product details
- ASIN : B0GTYKTZHG
- Publisher : O'Reilly Media
- Publication date : October 6, 2026
- Edition : 1st
- Language : English
- Print length : 376 pages
- ISBN-13 : 979-8341666931
- Item Weight : 1.32 pounds
- Dimensions : 7 x 0.78 x 9.19 inches
- Best Sellers Rank: #160,330 in Books (See Top 100 in Books)
About the author

Discover more of the author’s books, see similar authors, read book recommendations and more.











