Free Lesson

Vibe Check for PMs: Opus, Fable, GPT-5.6, Kimi K3

45 min
Aug 6, 2026 12:00 PM
Virtual (Zoom)

What you'll learn

What to test after you upgrade to a new model

Which model is actually best for PM tasks

The specific quirks of Opus, Fable 5, GPT-5.6 Sol and Kimi K

How to quickly create evals to benchmark new models

Why this topic matters

Four new frontier models shipped last month with stellar benchmarks, but how do they stand up to real-world usage? Which models work best as a thinking partner vs. a task executor? We test all four against real PM tasks live, show you how to prompt each one differently, and build an eval you can use to benchmark performance for your own tasks.

You'll learn from

Aman Khan

Aman Khan

AI PM at Arize, ex Spotify, Apple, Cruise

Eric Xiao

Eric Xiao

Founder @ Bloom, prev. AI PM at Meta, Arize

See all products from Aman