From Solo LLM to Agent Swarm: Scaling Multi-Agent AI with vLLM

About this session

As AI moves beyond single-model interactions, multi-agent systems are emerging as a powerful way to solve complex tasks through specialized, collaborating agents. This session explores how to build and serve multi-agent applications using open-source LLMs and vLLM, covering orchestration, agent specialization, parallel execution, validation, and efficient model serving. Through a practical demo, attendees will see how multiple agents can work together as an AI team to reason, verify, and deliver more reliable outcomes.

Speaker

Key takeaways

  • Design effective multi-agent architectures
  • Serve agents efficiently with vLLM
  • Build for reliability, not just demos

Related sessions