Cohort-based courses
Guided programs to get real results.
AI Evals For Engineers & PMs
4.7
·4 weeks·Sep 5 – Oct 3
Hamel Husain ML Engineer with 25 years of experience
Shreya Shankar ML Systems & Applied AI Evals Researcher
AI Evals and Analytics Playbook
4.9
·3 weeks·Aug 10 – Aug 31

Stella Liu Head of AI Applied Science
Amy Chen Cofounder, AI Evals & Analytics
Beyond Evals: Designing Improvement Flywheels for AI Products
4.8
·3 weeks·Sep 19 – Oct 10.png&w=256&q=75)

Aishwarya Naresh Reganti AI Founder & Advisor to F500s | Ex-AWS
Kiriti Badam Applied AI @ OpenAI Codex | Ex-Google
Build a Software Factory: Hands-off agentic coding for experienced engineers
4 weeks·Sep 15 – Oct 9


Matt Wynne Cucumber co-founder, BDD pioneer
+ Aldric Giacomoni, Zach Marcin, David Laing, and Jeremy Lightsmith
Production-Ready Systems with LLMs and Agents: An Intensive for Engineers
NEW·4 weeks·Oct 5 – Nov 2

Ehsan GazarStaff Engineer | AI & System Design
Build AI Agents for Enterprises
NEW·4 weeks·Aug 17 – Sep 13
Skanda VivekHow to build reliable agents
1-day workshops
Short, focused sessions to build specific skills.
Free Lightning Lessons
Interactive sessions to explore new topics.
Turn Eval Results Into a Better Model
·45 minutes403 StudentsWatch
Will Brown, Florian Brand, and Hamel HusainStop Paying Full Price for LLM Classification
·45 minutes362 StudentsWatch
Shreya Shankar and Hamel HusainHow to Setup Evals For Agents
·30 minutes2,560 StudentsWatch
Harrison Chase, Hamel Husain, andRaise Your Technical Bar as an AI-Native PM
·30 minutes16,029 StudentsWatch
Jason P. Yoong and Gayathri Keerthana (GK)AI Evals for Product Managers
·60 minutes2,084 StudentsWatch
Anshumani RuddraWhat Makes a Good Search Agent?
·60 minutes376 StudentsWatch
Nandan Thakur and Hamel HusainAI Evals for Building Reliable and Consistent Products
·60 minutes307 StudentsWatch
Anshumani RuddraDesign Evals Users Will Trust
·45 minutes802 StudentsWatch
Aishwarya Naresh RegantiProduction Grade AI Evals by Braintrust.dev
·30 minutes538 StudentsWatch
Mengying LiMastering Agentic RAG & AI Evals
·60 minutes1,335 StudentsWatch.png&w=1536&q=75)
Dr. Ryan Ahmed, Ph.D., MBA and Kukesh KodessModern Information Retrieval Evaluation In The RAG Era
·45 minutes5,406 StudentsWatch
Nandan Thakur, Hamel Husain, and Shreya ShankarDebug the weird stuff your AI does (in less than 1 hour)
·45 minutes5,188 StudentsWatch.webp&w=1536&q=75)
Marily Nika and Hamel HusainHow Evals Made GitHub Copilot Happen
·30 minutes901 StudentsWatch
John Berryman, Shawn Simister, and Hamel HusainLearn Agentic AI: Setting agents metrics and evaluations
·45 minutes879 StudentsWatch
Mahesh YadavBuild Your AI Evals & Analytics Playbook
·30 minutes558 StudentsWatch
Stella Liu and Amy ChenEvals in Action With Arize
·45 minutes224 StudentsWatch
Laurie VossPart 3: Building Robust Evaluations for AI Agents
·60 minutes158 StudentsWatch
Hamza Farooq and Gabriela de QueirozBuild Multi-Agent Systems You Can Audit
·30 minutes132 StudentsWatch
Stefan JansenEvals for Everyone
·3 lessons2,208 StudentsWatch
Kiriti & AishEvaluating AI Agents
·45 minutes1,449 StudentsWatch
Amir Feizpour and Samuel Dion-GirardeauPressure-test any AI analysis
·60 minutes842 StudentsWatch
Shane Butler, Sravya Madipalli, and Hai GuanPractical Evaluation Strategies for AI Agents
·45 minutes488 StudentsWatch
Hamza Farooq and Gabriela de QueirozFrom trading idea to validated strategy
·30 minutes125 StudentsWatch
Stefan JansenAutomating Evals With Claude Code + Phoenix
·60 minutes2,375 StudentsWatch
Mikyo King and Hamel HusainEvaluating Agentic AI Applications Beyond Vibe Checks
·45 minutes1,261 StudentsWatch
Aishwarya Naresh Reganti, Kiriti Badam, and Claire LongoUnderstanding Embedding Performance through Generative Evals
·60 minutes1,184 StudentsWatch
Jason Liu and Kelly HongHow OpenAI Customers Use Evals To Build Better AI Products
·30 minutes1,091 StudentsWatch
Jim Blomo and Hamel HusainBuild Your Own Eval Tools With Notebooks!
·45 minutes626 StudentsWatch
Vincent D. Warmerdam, Hamel Husain, and Shreya ShankarEvaluation Driven Development for Agentic AI Systems
·45 minutes598 StudentsWatch
Aurimas GriciūnasHow You Catch Production Hallucinations in Real Time
·60 minutes510 StudentsWatch
Jason Liu and Julia NeaguHow to Drive AI Evals Adoption
·30 minutes338 StudentsWatch
Dr Sebastian FoxSynthetic RAG evaluation
·60 minutes221 StudentsWatch
Alexey Grigorev and Doug TurnbullExperimentation in the AI Era: Lessons from the Trenches
·60 minutes132 StudentsWatch
Mirza Rahim Baig and Vishnukant PeddawadCollaborative AI Evals with Human Feedback
·30 minutes130 StudentsWatch
Rogério ChavesLLM-as-Judge: Grade Your AI Feature's Quality
·30 minutes24 StudentsWatch
Aki Wijesundara and Manu JayawardanaStop RAG Hallucinations With Citations
·30 minutes12 StudentsWatch
Aki Wijesundara and Manu JayawardanaError Analysis: The AI Engineer’s Best ROI
·60 minutes1,527 StudentsWatch
Hamel Husain and Shreya ShankarFrom Automation to Multi-Agent Architectures
·3 lessons1,368 StudentsWatch
Hamza FarooqSetting Eval for AI Agents & Scaling with Auto-Evaluation
·30 minutes874 StudentsWatch
Mahesh YadavOptimize Structured Data Retrieval With Evals
·45 minutes849 StudentsWatch
Daniel Svonava and Hamel HusainOnline Evals and Production Monitoring
·60 minutes834 StudentsWatch
Jason Liu, Ben Hylak, and Sidhant BendreEvaluate AI agents with Confidence
·45 minutes807 StudentsWatch
Mahesh YadavAI Systems Under Pressure: Red-Team Before You Ship
·60 minutes807 StudentsWatch
Krystal JacksonOptimize Your Dev Setup For Evals w/ Cursor Rules & MCP
·30 minutes693 StudentsWatch
Isaac Flath, Hamel Husain, and Shreya ShankarScaling Judge-Time Compute for Robust Auto LLM Evaluation
·60 minutes492 StudentsWatch
Jason Liu and Leonard TangStrategies for building self-improving document processing
·60 minutes433 StudentsWatch
Jason Liu and Eli BadgioMaster Evaluation Techniques for LLM Apps
·30 minutes417 StudentsWatch
Haroon ChouderyReliable RAG Agents: Intent-Driven Failure Detection
·60 minutes300 StudentsWatch
Jason Liu and Ben HylakCreate MCP Tool Evals Before You Ship
·45 minutes292 StudentsWatch
Emmanuel ParaskakisScale Evals Without the Chaos
·45 minutes263 StudentsWatch
Aishwarya Naresh RegantiEvals for Voice AI: Learnings from Google Evals Team
·30 minutes258 StudentsWatch
Ravin KumarMastering LLM Application Testing
·30 minutes245 StudentsWatch
Hugo Bowne-Anderson and Stefan KrawczykCalibrate LLM-as-a-judge for Real-world Impact
·45 minutes213 StudentsWatch
Eddie Landesberg🛠 Synthetic Data Flywheels: Build Reliable LLM Apps Faster
·30 minutes190 StudentsWatch
Hugo Bowne-Anderson and Stefan KrawczykDe-Risking LLM Model Switches w Evals & Prompt Optimization
·45 minutes147 StudentsWatch
Amir Feizpour and Hugo MailhotEvaluating AI Agents before Users Break Them
·60 minutes97 StudentsWatch
Aki Wijesundara, PhD, Marc Klingen, and Lotte VerheydenRun Eval Loops and Guardrails for Cursor Agents
·30 minutes97 StudentsWatch
Carmelo IariaStay Ahead in AI: Evaluate Any New LLM in 15 Minutes
·30 minutes96 StudentsWatch
Sherveen MashayekhiSetting up your first AI eval with a LLM-as-judge
·45 minutes74 StudentsWatch
Madalina Turlea and Catalina TurleaGo Beyond AI Evals: Diagnose and Decide
·45 minutes62 StudentsWatch
Rajiv ShahDebug Cursor Agent Failures Before Production
·30 minutes49 StudentsWatch
Carmelo IariaEvaluate Your RAG: Is It Actually Right?
·30 minutes18 StudentsWatch
Aki Wijesundara and Manu JayawardanaImprove reliability of your AI applications
·30 minutes747 StudentsWatch
Shreya RajpalUnderstand SHAP (SHapley Additive exPlanations)
·30 minutes313 StudentsWatch
Patrick HallDon't Tweak Prompts. Engineer Agents.
·30 minutes276 StudentsWatch
Hugo Bowne-Anderson and Skylar PayneShip a Production Cursor Agent System in 30 Minutes
·30 minutes221 StudentsWatch
Carmelo IariaThe Hidden Signal in Production AI Logs
·60 minutes176 StudentsWatch
Jason Liu and Scott ClarkHow to test and improve your AI agents
·45 minutes168 StudentsWatch
Jacob BankThe New Frontier of AI Search
·75 minutes124 StudentsWatch
Trey Grainger and Doug TurnbullHow to test AI when you don't have any data yet
·45 minutes27 StudentsWatch
Madalina Turlea and Catalina Turlea
