Observing and improving agents end-to-end (Langfuse, observability) across the idea-to-ship lifecycle, with autoresearch and CLI tooling.
Accessible with the Engineering + Workshops pass and above.
Getting an agent into production takes more than a good prompt: it needs somewhere to run code, credentials it can't leak, sessions that survive interruption, and infrastructure that scales. This talk traces how Anthropic's agentic surfaces evolved from the raw API to Claude Managed Agents, and what our Applied AI team has learned about harness design along the way.