Staging environment
Free Lesson

Evals in Action With Arize

Part of Evals for Everyone

45 min
Feb 27, 2026 12:00 PM
Virtual (Zoom)

In this video

What you'll learn

Build your first LLM-as-a-Judge evaluator

Write an eval that detects hallucinations in <10 minutes using Arize Phoenix's templates and your own custom criteria.

Trace your AI system end-to-end

Add observability to any LLM application so you can see exactly what's happening at every step, from input to output.

Choose the right evaluator for each failure mode

Learn when to use code-based checks, LLM judges, or human annotations based on what you're trying to catch.

Why this topic matters

You've learned why evals matter and what to measure. Now you need to actually build them. Most teams get stuck here because the gap between "understanding evals" and "shipping evals" feels enormous. This hands-on session bridges that gap with live code, real tools, and templates you can steal. You'll leave with working evaluators, not just concepts.

You'll learn from

Laurie Voss

Laurie Voss

Head of DevRel at Arize, co-founder, npm Inc

See all products from Aish & Kiriti