Vibe Check for PMs: Opus, Fable, GPT-5.6, Kimi K3

Hosted by Aman Khan and Eric Xiao

Thu, Aug 6, 2026

4:00 PM UTC (45 minutes)

Virtual (Zoom)

Free to join

189 students

Invite your network

Go deeper with a course

Claude Code for Product Managers (w/ Fable)
Aman Khan and Eric Xiao
View syllabus

What you'll learn

What to test after you upgrade to a new model

Which model is actually best for PM tasks

The specific quirks of Opus, Fable 5, GPT-5.6 Sol and Kimi K

How to quickly create evals to benchmark new models

Why this topic matters

Four new frontier models shipped last month with stellar benchmarks, but how do they stand up to real-world usage? Which models work best as a thinking partner vs. a task executor? We test all four against real PM tasks live, show you how to prompt each one differently, and build an eval you can use to benchmark performance for your own tasks.

You'll learn from

Aman Khan

AI PM at Arize, ex Spotify, Apple, Cruise

Aman has worked as a product leader at Arize AI, Spotify, Cruise, Zipline, and Apple. Currently, Aman is lead PM at Google on the Agent Platform. At Google, Aman helps teams launch and improve their AI systems. He recently led a popular deeplearning.ai course on Evaluating AI Agents, and has been featured by Lenny's Newsletter to cover AI Product Management a number of times.

Eric Xiao

Founder @ Bloom, prev. AI PM at Meta, Arize

Eric is a founder and full stack builder. He is currently building Bloom, an investing assistant (100k+ downloads). He previously led product at an AI evals company, holds a patent for an AI shopping assistant, and was an executive for a series C AI startup. Eric uses AI to prototype, market, and ship new products from scratch every day.

See all products from Aman

Sign up to join this lesson

By continuing, you agree to Maven's Terms and Privacy Policy.