From Solo LLM to Agent Swarm: Scaling Multi-Agent AI with vLLM
About this session
As AI moves beyond single-model interactions, multi-agent systems are emerging as a powerful way to solve complex tasks through specialized, collaborating agents. This session explores how to build and serve multi-agent applications using open-source LLMs and vLLM, covering orchestration, agent specialization, parallel execution, validation, and efficient model serving. Through a practical demo, attendees will see how multiple agents can work together as an AI team to reason, verify, and deliver more reliable outcomes.
Speaker
Key takeaways
- Design effective multi-agent architectures
- Serve agents efficiently with vLLM
- Build for reliability, not just demos
Related sessions
- Beyond Prompting: Building Production-Grade Systems with Context and Loop Engineering
- Evolving from Modern Data Engineering to AI Readiness: An Enterprise Execution Blueprint Abstract
- How Multi-Agent Systems Enable Personalization at Scale
- From Dashboards to Agentic Decision Systems: Designing the Enterprise Intelligence Layer