Staging environment
Free Lesson

Mastering LLM Application Testing

Part of The AI Evaluation Handbook

30 min
Mar 24, 2025 7:00 PM
Virtual (Zoom)

In this video

What you'll learn

Write effective pytest cases for testing LLM outputs.

Build a framework for iterative improvement of LLM apps

Evaluate LLM results systematically to improve stability

Why this topic matters

LLM-powered applications require constant iteration and evaluation to ensure robust and reliable outputs. In this Lightning Lesson, you’ll learn how to apply pytest to systematically test and refine your LLM apps. We’ll cover how to identify and address failure modes, evaluate outputs, and build confidence in your app’s performance.

You'll learn from

Hugo Bowne-Anderson

Hugo Bowne-Anderson

Podcaster, Educator, DS & ML expert

Stefan Krawczyk

Stefan Krawczyk

13+years in MLOps: Ex-Stitch Fix, Ex-Nextdoor, Ex-LinkedIn

Previously at

LinkedIn
Yale University
Stitch Fix
Stanford University
New York University
See all products from Hugo & Stefan