Back to Videos

Building Production-Ready AI Agents with Claude Managed Infrastructure

YouTube

Isabella He from Anthropic introduces Claude Managed Agents, a production ready infrastructure designed to help developers build and ship AI agents significantly faster. The presentation charts the evolution from raw model access via the Messages API to the Agent SDK, and finally to the Managed Agents system, which abstracts away complex orchestration tasks like context management, scaling, and observability. By moving the agent loop server side and decoupling it from tool execution, Managed Agents reduce latency and allow for seamless session persistence even when a user disconnects their local machine. The presentation features a hands on workshop where a functional Site Reliability Engineering agent is constructed to investigate a simulated system incident. This demonstration shows how the agent can autonomously query metrics, logs, and deployment data to identify a database pool exhaustion issue caused by a specific code refactor. Beyond basic task execution, the session highlights advanced capabilities such as subagents, long term memory through dreaming, and secure credential management via vaults. These tools transform agents from simple chatbots into autonomous teammates capable of handling sophisticated, high stakes engineering workflows.

Visual Summary

Infographic visualizing Building Production-Ready AI Agents with Claude Managed Infrastructure

This video covers the technical architecture and practical implementation of Claude Managed Agents, providing a roadmap for developers to transition from basic AI prototypes to production ready autonomous systems. Isabella He explains how this infrastructure simplifies the creation of AI agents by managing the underlying complexities of scaling, security, and orchestration. ## Key Takeaways * Managed Agents allow developers to build AI products 10 to 15 times faster by providing a purpose built harness. * The architecture decouples the agent loop (the brain) from tool execution (the hands), resulting in a 90 percent reduction in time to first token. * Systems are built on three primary resources: Agents for persona, Environments for infrastructure, and Sessions for user interaction. * Managed Agents handle complex primitives like context compaction, caching, and server side session persistence. * Advanced features like dreaming enable agents to self improve by reflecting on their own memory and logs. ## The Evolution of Agentic Interfaces The journey toward Managed Agents began with the Messages API, which provided raw access to Claude models but required developers to manually handle context management and agent loops. This was followed by the Agent SDK, which introduced more powerful orchestration but still left the burden of hosting and scaling on the developer. Managed Agents represent the next step, offering a server side loop that runs on Anthropic managed infrastructure. This shift means that even if a developer closes their laptop or a local container crashes, the agent loop continues to run reliably. This architecture is specifically designed to combat context anxiety, a behavior where models wrap up tasks too early when they feel the context window is becoming crowded. The managed harness includes built in mitigations to ensure Claude performs optimally regardless of task length. ## The Three Pillars of Managed Agents Building a Managed Agent requires defining three distinct resources. The first is the Agent endpoint, which defines the persona, model choice, system prompts, and specific skills. This is effectively the brain of the operation. The second is the Environment, which establishes where the agent runs, including networking configurations and security guardrails. This serves as the hands of the agent, providing a sandboxed space to execute actions. Finally, the Session ties these two together, creating a unique instance of interaction that is persisted in the cloud. Because these sessions are managed server side, they support complex events and webhooks, allowing agents to respond to external triggers without requiring constant active polling from a client. ## Architecture: Decoupling Brain and Hands A major architectural innovation in Managed Agents is the decoupling of the agent loop from the tool execution sandbox. Previously, these components were often tightly coupled in a single container. By separating them, Anthropic has achieved significant performance gains. The agent loop is a single service managing thousands of sessions, while sandboxes are provisioned on demand only when a tool actually needs to be executed. This approach not only improves security by isolating credentials but also dramatically reduces latency. Developers see a massive improvement in the responsiveness of their agents, as the system no longer needs to spin up a heavy container at the start of every single user session. ## Practical Applications The most compelling demonstration of this technology is the creation of a Site Reliability Engineering (SRE) agent. In high pressure engineering environments, investigating a 3 AM system outage is a tedious and error prone task for humans. A Managed Agent can be programmed to autonomously triage incidents by fetching metrics, analyzing logs, and identifying problematic code commits. By providing the agent with access to local tools through a streaming protocol, it can perform deep dives into system health and suggest remediation steps, such as rolling back a specific deployment. This reduces the burden on human on call rotations and ensures faster mean time to resolution for critical bugs. ## Frequently Asked Questions ### How does session persistence work if my local app crashes? Because the agent loop runs on Anthropic server side infrastructure, the session state is maintained independently of your local client. When your application reconnects and requests the session ID, it can resume exactly where it left off, including all previous tool outputs and message history stored in the cloud. ### Can I use my own infrastructure for tool execution? Yes, Managed Agents now support bringing your own compute. This allows you to execute the hands of the agent within your own VPC or local environment while still leveraging the managed brain loop in the Anthropic cloud. This is ideal for organizations with strict data residency or security requirements. ### What is the difference between memory and dreaming? Memory refers to the persistent storage of session data, while dreaming is an advanced service that allows agents to periodically review their own logs. During a dream, the agent can identify patterns, learn from user corrections, and decide which information is worth retaining for future sessions, effectively self improving its own performance over time.

Diagram

Loading diagram...

Timestamps

00:00
IntroductionIsabella He introduces herself and the Applied AI team at Anthropic.
02:09
Evolution of Agent InterfacesTransition from Messages API to the Agent SDK and finally to Managed Agents.
04:31
Benefits of Managed AgentsOverview of speed to production and infrastructure abstraction.
05:54
The Three Primary ResourcesDefining Agents, Environments, and Sessions.
07:22
Decoupled ArchitectureExplaining the performance gains from separating the loop and the sandbox.
11:10
SRE Agent WorkshopHands-on demonstration of building an incident response investigator.
23:35
Advanced CapabilitiesOverview of subagents, memory, dreaming, and secure vaults.

Target Audience

Software engineers, AI developers, and Site Reliability Engineers looking to build production grade autonomous AI systems.

Use Cases

  • -Automating 24/7 incident response for complex cloud environments.
  • -Building customer support agents with persistent memory and secure data access.
  • -Creating autonomous coding assistants that can interact with local files and remote APIs.
  • -Developing data analysis agents that manage their own scaling and infrastructure.

Key Topics

Managed AI InfrastructureAgentic Workflow EvolutionServer-side Agent LoopsAutonomous Incident ResponseAdvanced Agent Capabilities