Agentic AI Notebook
Back to Roadmap
Phase 19

Eval Engineering & Observability

How to prove an agent works: datasets, LLM-as-judge, online/offline eval, benchmarks, agent-specific OpenTelemetry, and regression suites.

18 modules

Estimated read time is shown per module based on lesson length.