Free Lesson
Evals in Action With Arize
Part of Evals for Everyone
45 min
Feb 27, 2026 12:00 PM
Virtual (Zoom)
In this video
What you'll learn
Build your first LLM-as-a-Judge evaluator
Write an eval that detects hallucinations in <10 minutes using Arize Phoenix's templates and your own custom criteria.
Trace your AI system end-to-end
Add observability to any LLM application so you can see exactly what's happening at every step, from input to output.
Choose the right evaluator for each failure mode
Learn when to use code-based checks, LLM judges, or human annotations based on what you're trying to catch.
Why this topic matters
You've learned why evals matter and what to measure. Now you need to actually build them. Most teams get stuck here because the gap between "understanding evals" and "shipping evals" feels enormous. This hands-on session bridges that gap with live code, real tools, and templates you can steal. You'll leave with working evaluators, not just concepts.
You'll learn from

Laurie Voss
Head of DevRel at Arize, co-founder, npm Inc
