All courses
33 published courses. Narrow by subject, role, level or format, or search by the outcome you are after.
Filter1
Filter courses
DoneSubject
For your role
AI for Engineers3AI for Product Managers4AI for Leaders3AI for Designers0AI for Marketers0AI for Data & Analysts5AI for Founders2
Competency1
Format
Matching courses
CourseRating
Technical AILLM-as-Judge Lab: Calibration and DriftA hands-on workshop for data analysts who already call LLM APIs and know eval-framework basics.Workshop · 1.4 hoursNot yet rated14 lessons
Technical AIBuilding an Eval HarnessFor AI-savvy data analysts moving from ad hoc prompt testing to a real evaluation practice, this course builds a working eval harness end to end: a labeled test dataset, deterministic and LLM-as-judge graders, pass/fail thresholds tied to real stakes, regression detection, and a CI-wired scorecard.Standard (2–4 h) · 2 hoursNot yet rated22 lessons
AI for BusinessHow to Tell If Your AI Is Any GoodFor non-technical leaders who sponsor or review AI products, this course cuts through eval jargon and vendor hype.Quick (30–60 min) · 45 minNot yet rated11 lessons
AI LiteracyLLM Guardrails, Evals and Agentic MemoryLearn to take LLM applications from prototype to production.Quick (30–60 min) · 2.9 hoursNot yet rated22 lessons
Technical AIBuilding Your First Production AgentA hands-on course for engineers who already call LLM APIs and want to ship a tool-using agent that survives production: recovering from failed tool calls, capping runaway spend, extending itself with an MCP server, passing task-completion evals, and deploying with logging enabled.Quick (30–60 min) · 1.8 hoursNot yet rated23 lessons
AI for BusinessPrompt Patterns for Product DiscoveryAn intermediate, no-code course for product managers who already prompt chat-based LLMs and know discovery frameworks like JTBD.Quick (30–60 min) · 1.4 hoursNot yet rated19 lessons
AI GovernanceWhat an Agent Actually Is (and What It Cannot Do)A foundational course for non-technical leaders who need to tell a genuine agentic system apart from a relabeled chatbot, recognize the concrete ways agents fail in production, judge when a task is too risky or irreversible to hand to an agent, and interrogate a vendor's 'agent' claim before sign…Quick (30–60 min) · 50 minNot yet rated15 lessons
AI LiteracyEvaluating LLM Output QualityHow to tell whether an LLM feature is actually working.Quick (30–60 min) · 3.4 hoursNot yet rated16 lessons
Matching modules
24 modules- Module 3Failure-Resilient Agent Loopsin Building Your First Production Agent · 4 lessonsai-for-engineersagentic-aiai-evalsworkflowsquickintermediate
- Module 6Evaluation for Task Completionin Building Your First Production Agent · 3 lessonsai-for-engineersai-for-data-analystsagentic-aiai-evalsquickintermediate
- Module 1Start Herein Building an Eval Harness · 1 lessonai-for-data-analystsai-evalsstandardintermediate
- Module 2Foundations of Eval Harnessesin Building an Eval Harness · 4 lessonsai-for-data-analystsai-evalsstandardintermediate
- Module 3Building the Test Datasetin Building an Eval Harness · 3 lessonsai-for-data-analystsai-evalsstandardintermediate
- Module 4Writing Graders and Scoring Rubricsin Building an Eval Harness · 4 lessonsai-for-data-analystsai-evalsstandardintermediate
- Module 5Thresholds, Regressions, and Comparisonsin Building an Eval Harness · 4 lessonsai-for-data-analystsai-evalsstandardintermediate
- Module 6Diagnosing Failures and Operationalizing the Harnessin Building an Eval Harness · 6 lessonsai-for-data-analystsai-evalsstandardintermediate
- Module 1Why LLM Outputs Break: Understanding Prompt Driftin Evaluating LLM Output Quality · 3 lessonsai-for-engineersai-for-product-managersai-for-data-analystsprompt-engineeringai-evalsquickfoundational
- Module 2Structured Output as a Contractin Evaluating LLM Output Quality · 3 lessonsai-for-engineersai-for-data-analystsprompt-engineeringai-evalsquickfoundational
- Module 3Corrective Retries: Recovering from Validation Failuresin Evaluating LLM Output Quality · 3 lessonsai-for-engineersai-for-data-analystsprompt-engineeringai-evalsworkflowsquickintermediate
- Module 4Measuring What Matters: Evaluation Metricsin Evaluating LLM Output Quality · 4 lessonsai-for-engineersai-for-product-managersai-for-data-analystsprompt-engineeringai-evalsquickintermediate
- Module 5Putting It Together: A Production-Ready Evaluation Pipelinein Evaluating LLM Output Quality · 3 lessonsai-for-engineersai-for-product-managersai-for-data-analystsprompt-engineeringai-evalsworkflowsquickadvanced
- Module 1Start Herein How to Tell If Your AI Is Any Good · 1 lessonai-for-leadersai-evalsquickfoundational
- Module 2Reading the Benchmark Landscapein How to Tell If Your AI Is Any Good · 3 lessonsai-for-leadersai-evalsquickfoundational
- Module 3From Score to Production Signalin How to Tell If Your AI Is Any Good · 3 lessonsai-for-leadersai-evalsquickfoundational
- Module 4Interrogating the Vendorin How to Tell If Your AI Is Any Good · 4 lessonsai-for-leadersai-evalsquickfoundational
- Module 2LLM Evaluation (Evals): Building Rigorous Testing Pipelinesin LLM Guardrails, Evals and Agentic Memory · 4 lessonsai-for-engineersai-for-data-analystscontext-ragai-evalsquickintermediate
- Module 1Start Herein LLM-as-Judge Lab: Calibration and Drift · 1 lessonai-for-data-analystsai-evalsworkshopadvanced
- Module 2Designing and Implementing an LLM Judgein LLM-as-Judge Lab: Calibration and Drift · 4 lessonsai-for-data-analystsai-evalsworkshopadvanced
- Module 3Calibrating the Judge Against Human Ratersin LLM-as-Judge Lab: Calibration and Drift · 3 lessonsai-for-data-analystsai-evalsworkshopadvanced
- Module 4Monitoring Drift and Comparing Judgesin LLM-as-Judge Lab: Calibration and Drift · 6 lessonsai-for-data-analystsai-evalsworkshopadvanced
- Module 4Evaluating AI-Generated Personasin Prompt Patterns for Product Discovery · 3 lessonsai-for-product-managersprompt-engineeringai-evalsresponsible-aiquickintermediate
- Module 3How Agents Fail in Productionin What an Agent Actually Is (and What It Cannot Do) · 4 lessonsai-for-product-managersai-for-leadersai-for-foundersagentic-aiai-evalsresponsible-aiquickfoundational