The Hidden Distributed Systems Problem Behind Agentic AI
About this session
Agentic AI becomes a distributed systems problem once multiple agents begin sharing state, calling tools, exchanging messages, and executing asynchronous workflows. In this lightning talk, Jay Dave examines common production failure modes including duplicate execution, stale state, failed handoffs, retry loops, and limited visibility into agent behavior. He will introduce practical engineering patterns for orchestration, state ownership, failure isolation, observability, and recovery that help teams move multi-agent systems from impressive prototypes to reliable production platforms.
Speaker
Key takeaways
- Identify distributed systems failure modes that emerge in multi-agent architectures.
- Use clearer orchestration, state ownership, and retry patterns to improve reliability.
- Design observability and recovery paths before deploying autonomous workflows into production.