Cohort-based courses
Guided programs to get real results.
AI Evals For Engineers & PMs
4.7
·6 weeks·Oct 10 – Nov 21
Hamel Husain ML Engineer with 25 years of experience
Shreya Shankar ML Systems & Applied AI Evals Researcher
Forward Deployed AI Engineer Bootcamp
4.4
·4 weeks·Oct 20 – Nov 13

Meri Nova AI Educator to 145k Linkedin
Hai Nghiem Senior AI Engineer, Investor & Advisor
AI Evals Certification: Master Building Reliable AI
4.7


Stella Liu Head of AI Applied Science
Amy Chen Cofounder, AI Evals & Analytics
Beyond Evals: Designing Improvement Flywheels for AI Products
4.8
.png&w=256&q=75)

Aishwarya Naresh Reganti AI Founder & Advisor to F500s | Ex-AWS
Kiriti Badam Applied AI @ OpenAI Codex | Ex-Google
Production-Ready Systems with LLMs and Agents: An Intensive for Engineers
5.0
·3 weeks·Nov 30 – Dec 22
Ehsan GazarStaff Engineer | AI & System Design
Build AI Agents for Enterprises
4 weeks·Oct 5 – Nov 1
Skanda VivekHow to build reliable agents
1-day workshops
Short, focused sessions to build specific skills.
Free Lightning Lessons
Interactive sessions to explore new topics.
Raise Your Technical Bar as an AI-Native PM
·30 minutes16,137 StudentsWatch
Jason P. Yoong and Gayathri Keerthana (GK)Modern Information Retrieval Evaluation In The RAG Era
·45 minutes5,427 StudentsWatch
Nandan Thakur, Hamel Husain, and Shreya ShankarDebug the weird stuff your AI does (in less than 1 hour)
·45 minutes5,201 StudentsWatch.webp&w=1536&q=75)
Marily Nika and Hamel HusainHow to Setup Evals For Agents
·30 minutes2,711 StudentsWatch
Harrison Chase, Hamel Husain, andAI Evals for Product Managers
·60 minutes2,622 StudentsWatch
Anshumani RuddraAutomating Evals With Claude Code + Phoenix
·60 minutes2,383 StudentsWatch
Mikyo King and Hamel HusainEvals for Everyone
·3 lessons2,211 StudentsWatch
Kiriti & AishError Analysis: The AI Engineer’s Best ROI
·60 minutes1,539 StudentsWatch
Hamel Husain and Shreya ShankarEvaluating AI Agents
·45 minutes1,461 StudentsWatch
Amir Feizpour and Samuel Dion-GirardeauFrom Automation to Multi-Agent Architectures
·3 lessons1,369 StudentsWatch
Hamza FarooqMastering Agentic RAG & AI Evals
·60 minutes1,362 StudentsWatch.png&w=1536&q=75)
Dr. Ryan Ahmed, Ph.D., MBA and Kukesh KodessEvaluating Agentic AI Applications Beyond Vibe Checks
·45 minutes1,264 StudentsWatch
Aishwarya Naresh Reganti, Kiriti Badam, and Claire LongoUnderstanding Embedding Performance through Generative Evals
·60 minutes1,187 StudentsWatch
Jason Liu and Kelly HongHow OpenAI Customers Use Evals To Build Better AI Products
·30 minutes1,099 StudentsWatch
Jim Blomo and Hamel HusainHow Evals Made GitHub Copilot Happen
·30 minutes909 StudentsWatch
John Berryman, Shawn Simister, and Hamel HusainLearn Agentic AI: Setting agents metrics and evaluations
·45 minutes887 StudentsWatch
Mahesh YadavSetting Eval for AI Agents & Scaling with Auto-Evaluation
·30 minutes881 StudentsWatch
Mahesh YadavOptimize Structured Data Retrieval With Evals
·45 minutes851 StudentsWatch
Daniel Svonava and Hamel HusainPressure-test any AI analysis
·60 minutes846 StudentsWatch
Shane Butler, Sravya Madipalli, and Hai GuanOnline Evals and Production Monitoring
·60 minutes836 StudentsWatch
Jason Liu, Ben Hylak, and Sidhant BendreDesign Evals Users Will Trust
·45 minutes812 StudentsWatch
Aishwarya Naresh RegantiEvaluate AI agents with Confidence
·45 minutes811 StudentsWatch
Mahesh YadavAI Systems Under Pressure: Red-Team Before You Ship
·60 minutes808 StudentsWatch
Krystal JacksonImprove reliability of your AI applications
·30 minutes747 StudentsWatch
Shreya RajpalOptimize Your Dev Setup For Evals w/ Cursor Rules & MCP
·30 minutes697 StudentsWatch
Isaac Flath, Hamel Husain, and Shreya ShankarBuild Your Own Eval Tools With Notebooks!
·45 minutes630 StudentsWatch
Vincent D. Warmerdam, Hamel Husain, and Shreya ShankarEvaluation Driven Development for Agentic AI Systems
·45 minutes601 StudentsWatch
Aurimas GriciūnasBuild Your AI Evals & Analytics Playbook
·30 minutes575 StudentsWatch
Stella Liu and Amy ChenProduction Grade AI Evals by Braintrust.dev
·30 minutes555 StudentsWatch
Mengying LiHow You Catch Production Hallucinations in Real Time
·60 minutes514 StudentsWatch
Jason Liu and Julia NeaguTurn Eval Results Into a Better Model
·45 minutes511 StudentsWatch
Will Brown, Florian Brand, and Hamel HusainPractical Evaluation Strategies for AI Agents
·45 minutes494 StudentsWatch
Hamza Farooq and Gabriela de QueirozScaling Judge-Time Compute for Robust Auto LLM Evaluation
·60 minutes492 StudentsWatch
Jason Liu and Leonard TangStrategies for building self-improving document processing
·60 minutes434 StudentsWatch
Jason Liu and Eli BadgioMaster Evaluation Techniques for LLM Apps
·30 minutes419 StudentsWatch
Haroon ChouderyWhat Makes a Good Search Agent?
·60 minutes418 StudentsWatch
Nandan Thakur and Hamel HusainStop Paying Full Price for LLM Classification
·45 minutes413 StudentsWatch
Shreya Shankar and Hamel HusainHow to Drive AI Evals Adoption
·30 minutes348 StudentsWatch
Dr Sebastian FoxAI Evals for Building Reliable and Consistent Products
·60 minutes330 StudentsWatch
Anshumani RuddraUnderstand SHAP (SHapley Additive exPlanations)
·30 minutes314 StudentsWatch
Patrick HallReliable RAG Agents: Intent-Driven Failure Detection
·60 minutes301 StudentsWatch
Jason Liu and Ben HylakCreate MCP Tool Evals Before You Ship
·45 minutes296 StudentsWatch
Emmanuel ParaskakisDon't Tweak Prompts. Engineer Agents.
·30 minutes277 StudentsWatch
Hugo Bowne-Anderson and Skylar PayneEvals for Voice AI: Learnings from Google Evals Team
·30 minutes268 StudentsWatch
Ravin KumarScale Evals Without the Chaos
·45 minutes263 StudentsWatch
Aishwarya Naresh RegantiMastering LLM Application Testing
·30 minutes247 StudentsWatch
Hugo Bowne-Anderson and Stefan KrawczykEvals in Action With Arize
·45 minutes229 StudentsWatch
Laurie VossSynthetic RAG evaluation
·60 minutes224 StudentsWatch
Alexey Grigorev and Doug TurnbullShip a Production Cursor Agent System in 30 Minutes
·30 minutes223 StudentsWatch
Carmelo IariaCalibrate LLM-as-a-judge for Real-world Impact
·45 minutes218 StudentsWatch
Eddie Landesberg🛠 Synthetic Data Flywheels: Build Reliable LLM Apps Faster
·30 minutes190 StudentsWatch
Hugo Bowne-Anderson and Stefan KrawczykThe Hidden Signal in Production AI Logs
·60 minutes177 StudentsWatch
Jason Liu and Scott ClarkHow to test and improve your AI agents
·45 minutes174 StudentsWatch
Jacob BankPart 3: Building Robust Evaluations for AI Agents
·60 minutes169 StudentsWatch
Hamza Farooq and Gabriela de QueirozDe-Risking LLM Model Switches w Evals & Prompt Optimization
·45 minutes147 StudentsWatch
Amir Feizpour and Hugo MailhotThe New Frontier of AI Search
·75 minutes141 StudentsWatch
Trey Grainger and Doug TurnbullCollaborative AI Evals with Human Feedback
·30 minutes137 StudentsWatch
Rogério ChavesBuild Multi-Agent Systems You Can Audit
·30 minutes133 StudentsWatch
Stefan JansenExperimentation in the AI Era: Lessons from the Trenches
·60 minutes133 StudentsWatch
Mirza Rahim Baig and Vishnukant PeddawadFrom trading idea to validated strategy
·30 minutes130 StudentsWatch
Stefan JansenRun Eval Loops and Guardrails for Cursor Agents
·30 minutes99 StudentsWatch
Carmelo IariaStay Ahead in AI: Evaluate Any New LLM in 15 Minutes
·30 minutes97 StudentsWatch
Sherveen MashayekhiSetting up your first AI eval with a LLM-as-judge
·45 minutes82 StudentsWatch
Madalina Turlea and Catalina TurleaGo Beyond AI Evals: Diagnose and Decide
·45 minutes65 StudentsWatch
Rajiv ShahDebug Cursor Agent Failures Before Production
·30 minutes49 StudentsWatch
Carmelo IariaHow to test AI when you don't have any data yet
·45 minutes29 StudentsWatch
Madalina Turlea and Catalina Turlea

