Replace vibe-checks with real data for your AI products
About this session
Whether you're building your first AI feature or exploring what AI could do in your product, there's a moment every PM hits: you look at the AI output and think "this seems okay?" but you don't have a structured way to know for sure.
You've heard about evals. You know vibe checking isn't enough. But where do you actually start?
This workshop takes you from that starting point to a working evaluation suit, built on a real AI feature use case
Speaker
Key takeaways
- Understand the most important components of building AI features at scale
- Build an evaluation suite to scale your testing
- Learn how to identify failure patterns. Then iterate on both prompt and evals