Free Lesson
Vibe Check for PMs: Opus, Fable, GPT-5.6, Kimi K3
45 min
Aug 6, 2026 12:00 PM
Virtual (Zoom)
What you'll learn
What to test after you upgrade to a new model
Which model is actually best for PM tasks
The specific quirks of Opus, Fable 5, GPT-5.6 Sol and Kimi K
How to quickly create evals to benchmark new models
Why this topic matters
Four new frontier models shipped last month with stellar benchmarks, but how do they stand up to real-world usage? Which models work best as a thinking partner vs. a task executor? We test all four against real PM tasks live, show you how to prompt each one differently, and build an eval you can use to benchmark performance for your own tasks.
You'll learn from

Aman Khan
AI PM at Arize, ex Spotify, Apple, Cruise

Eric Xiao
Founder @ Bloom, prev. AI PM at Meta, Arize
.png&w=1536&q=75)