Free Courses
Hands-on, first-principles micro-courses on designing, evaluating, and deploying generative AI systems.
Featured Course
Autoraters 101: The Crash Course
The complete guide to LLM-as-a-judge evaluation systems. Stop relying on subjective manual reviews or broken traditional metrics. Learn why larger models can objectively rate smaller production pipelines, calibrate rubrics against human reference sets, and run repeatable automated evals.
Upcoming Tracks
Production LLM Monitoring & Drift
Calibrating autoraters in CI/CD pipelines, tracking statistical drift over time, and catching subtle model degradations before users do.
RAG Blueprints & Grounding
Building hallucination-free retrieval pipelines, deterministic context assembly, and automated factuality filters for enterprise data.
Never miss a new release
Get new video mini-courses, deep-dive evaluation playbooks, and runnable Python templates delivered straight to your inbox.
Subscribe — It's Free