Cohort-based courses
Guided programs to get real results.
AI Evals For Engineers & PMs
4.7
·4 weeks·Sep 5 – Oct 3
Hamel Husain ML Engineer with 25 years of experience
Shreya Shankar ML Systems & Applied AI Evals Researcher
AI Evals Certification: Master Building Reliable AI
4.8
·3 weeks·Aug 31 – Sep 21

Stella Liu Head of AI Applied Science
Amy Chen Cofounder, AI Evals & Analytics
Beyond Evals: Designing Improvement Flywheels for AI Products
4.8
·3 weeks·Sep 19 – Oct 10.png&w=256&q=75)

Aishwarya Naresh Reganti AI Founder & Advisor to F500s | Ex-AWS
Kiriti Badam Applied AI @ OpenAI Codex | Ex-Google
Build a Software Factory: Hands-off agentic coding for experienced engineers
4.5
·4 weeks·Sep 15 – Oct 9

Matt Wynne Cucumber co-founder, BDD pioneer
+ Aldric Giacomoni, David Laing, and Jeremy Lightsmith
Production-Ready Systems with LLMs and Agents: An Intensive for Engineers
5.0
·4 weeks·Oct 5 – Nov 2
Ehsan GazarStaff Engineer | AI & System Design
Build AI Agents for Enterprises
4.0
·4 weeks·Oct 5 – Nov 1Skanda VivekHow to build reliable agents
1-day workshops
Short, focused sessions to build specific skills.
Free Lightning Lessons
Interactive sessions to explore new topics.
Turn Eval Results Into a Better Model
·45 minutes473 StudentsWatch
Will Brown, Florian Brand, and Hamel HusainHow to Setup Evals For Agents
·30 minutes2,645 StudentsWatch
Harrison Chase, Hamel Husain, andRaise Your Technical Bar as an AI-Native PM
·30 minutes16,079 StudentsWatch
Jason P. Yoong and Gayathri Keerthana (GK)What Makes a Good Search Agent?
·60 minutes405 StudentsWatch
Nandan Thakur and Hamel HusainThe New Frontier of AI Search
·75 minutes134 StudentsWatch
Trey Grainger and Doug TurnbullAI Evals for Product Managers
·60 minutes2,099 StudentsWatch
Anshumani RuddraStop Paying Full Price for LLM Classification
·45 minutes400 StudentsWatch
Shreya Shankar and Hamel HusainMastering Agentic RAG & AI Evals
·60 minutes1,347 StudentsWatch.png&w=1536&q=75)
Dr. Ryan Ahmed, Ph.D., MBA and Kukesh KodessHow Evals Made GitHub Copilot Happen
·30 minutes905 StudentsWatch
John Berryman, Shawn Simister, and Hamel HusainProduction Grade AI Evals by Braintrust.dev
·30 minutes545 StudentsWatch
Mengying LiHow to Drive AI Evals Adoption
·30 minutes342 StudentsWatch
Dr Sebastian FoxModern Information Retrieval Evaluation In The RAG Era
·45 minutes5,413 StudentsWatch
Nandan Thakur, Hamel Husain, and Shreya ShankarError Analysis: The AI Engineer’s Best ROI
·60 minutes1,529 StudentsWatch
Hamel Husain and Shreya ShankarEvaluating AI Agents
·45 minutes1,453 StudentsWatch
Amir Feizpour and Samuel Dion-GirardeauHow OpenAI Customers Use Evals To Build Better AI Products
·30 minutes1,097 StudentsWatch
Jim Blomo and Hamel HusainLearn Agentic AI: Setting agents metrics and evaluations
·45 minutes882 StudentsWatch
Mahesh YadavEvaluation Driven Development for Agentic AI Systems
·45 minutes601 StudentsWatch
Aurimas GriciūnasPractical Evaluation Strategies for AI Agents
·45 minutes492 StudentsWatch
Hamza Farooq and Gabriela de QueirozSynthetic RAG evaluation
·60 minutes223 StudentsWatch
Alexey Grigorev and Doug TurnbullPart 3: Building Robust Evaluations for AI Agents
·60 minutes163 StudentsWatch
Hamza Farooq and Gabriela de QueirozEvaluating AI Agents before Users Break Them
·60 minutes103 StudentsWatch
Aki Wijesundara, PhD, Marc Klingen, and Lotte VerheydenDebug the weird stuff your AI does (in less than 1 hour)
·45 minutes5,192 StudentsWatch.webp&w=1536&q=75)
Marily Nika and Hamel HusainAutomating Evals With Claude Code + Phoenix
·60 minutes2,380 StudentsWatch
Mikyo King and Hamel HusainUnderstanding Embedding Performance through Generative Evals
·60 minutes1,186 StudentsWatch
Jason Liu and Kelly HongSetting Eval for AI Agents & Scaling with Auto-Evaluation
·30 minutes878 StudentsWatch
Mahesh YadavOptimize Structured Data Retrieval With Evals
·45 minutes851 StudentsWatch
Daniel Svonava and Hamel HusainDesign Evals Users Will Trust
·45 minutes807 StudentsWatch
Aishwarya Naresh RegantiOptimize Your Dev Setup For Evals w/ Cursor Rules & MCP
·30 minutes697 StudentsWatch
Isaac Flath, Hamel Husain, and Shreya ShankarBuild Your Own Eval Tools With Notebooks!
·45 minutes628 StudentsWatch
Vincent D. Warmerdam, Hamel Husain, and Shreya ShankarBuild Your AI Evals & Analytics Playbook
·30 minutes565 StudentsWatch
Stella Liu and Amy ChenAI Evals for Building Reliable and Consistent Products
·60 minutes317 StudentsWatch
Anshumani RuddraCollaborative AI Evals with Human Feedback
·30 minutes132 StudentsWatch
Rogério ChavesFrom trading idea to validated strategy
·30 minutes127 StudentsWatch
Stefan JansenRun Eval Loops and Guardrails for Cursor Agents
·30 minutes99 StudentsWatch
Carmelo IariaSetting up your first AI eval with a LLM-as-judge
·45 minutes79 StudentsWatch
Madalina Turlea and Catalina TurleaEvals for Everyone
·3 lessons2,208 StudentsWatch
Kiriti & AishFrom Automation to Multi-Agent Architectures
·3 lessons1,368 StudentsWatch
Hamza FarooqEvaluating Agentic AI Applications Beyond Vibe Checks
·45 minutes1,263 StudentsWatch
Aishwarya Naresh Reganti, Kiriti Badam, and Claire LongoPressure-test any AI analysis
·60 minutes844 StudentsWatch
Shane Butler, Sravya Madipalli, and Hai GuanOnline Evals and Production Monitoring
·60 minutes835 StudentsWatch
Jason Liu, Ben Hylak, and Sidhant BendreEvaluate AI agents with Confidence
·45 minutes808 StudentsWatch
Mahesh YadavAI Systems Under Pressure: Red-Team Before You Ship
·60 minutes808 StudentsWatch
Krystal JacksonImprove reliability of your AI applications
·30 minutes747 StudentsWatch
Shreya RajpalHow You Catch Production Hallucinations in Real Time
·60 minutes512 StudentsWatch
Jason Liu and Julia NeaguScaling Judge-Time Compute for Robust Auto LLM Evaluation
·60 minutes492 StudentsWatch
Jason Liu and Leonard TangStrategies for building self-improving document processing
·60 minutes434 StudentsWatch
Jason Liu and Eli BadgioMaster Evaluation Techniques for LLM Apps
·30 minutes417 StudentsWatch
Haroon ChouderyUnderstand SHAP (SHapley Additive exPlanations)
·30 minutes314 StudentsWatch
Patrick HallReliable RAG Agents: Intent-Driven Failure Detection
·60 minutes300 StudentsWatch
Jason Liu and Ben HylakCreate MCP Tool Evals Before You Ship
·45 minutes293 StudentsWatch
Emmanuel ParaskakisDon't Tweak Prompts. Engineer Agents.
·30 minutes276 StudentsWatch
Hugo Bowne-Anderson and Skylar PayneScale Evals Without the Chaos
·45 minutes263 StudentsWatch
Aishwarya Naresh RegantiEvals for Voice AI: Learnings from Google Evals Team
·30 minutes261 StudentsWatch
Ravin KumarMastering LLM Application Testing
·30 minutes246 StudentsWatch
Hugo Bowne-Anderson and Stefan KrawczykEvals in Action With Arize
·45 minutes225 StudentsWatch
Laurie VossShip a Production Cursor Agent System in 30 Minutes
·30 minutes223 StudentsWatch
Carmelo IariaCalibrate LLM-as-a-judge for Real-world Impact
·45 minutes215 StudentsWatch
Eddie Landesberg🛠 Synthetic Data Flywheels: Build Reliable LLM Apps Faster
·30 minutes190 StudentsWatch
Hugo Bowne-Anderson and Stefan KrawczykThe Hidden Signal in Production AI Logs
·60 minutes176 StudentsWatch
Jason Liu and Scott ClarkHow to test and improve your AI agents
·45 minutes168 StudentsWatch
Jacob BankDe-Risking LLM Model Switches w Evals & Prompt Optimization
·45 minutes147 StudentsWatch
Amir Feizpour and Hugo MailhotBuild Multi-Agent Systems You Can Audit
·30 minutes132 StudentsWatch
Stefan JansenExperimentation in the AI Era: Lessons from the Trenches
·60 minutes132 StudentsWatch
Mirza Rahim Baig and Vishnukant PeddawadStay Ahead in AI: Evaluate Any New LLM in 15 Minutes
·30 minutes97 StudentsWatch
Sherveen MashayekhiGo Beyond AI Evals: Diagnose and Decide
·45 minutes64 StudentsWatch
Rajiv ShahDebug Cursor Agent Failures Before Production
·30 minutes49 StudentsWatch
Carmelo IariaHow to test AI when you don't have any data yet
·45 minutes28 StudentsWatch
Madalina Turlea and Catalina TurleaLLM-as-Judge: Grade Your AI Feature's Quality
·30 minutes25 StudentsWatch
Aki Wijesundara and Manu JayawardanaEvaluate Your RAG: Is It Actually Right?
·30 minutes19 StudentsWatch
Aki Wijesundara and Manu JayawardanaStop RAG Hallucinations With Citations
·30 minutes14 StudentsWatch
Aki Wijesundara and Manu Jayawardana

