October 2026
Intermediate to advanced
225 pages
4h 39m
English
Chapter 1: Introduction (available)
Chapter 2: LLMs and Evaluation Basics (available)
Chapter 3: Error Analysis (available)
Chapter 4: Collaborative Evaluation Practices (available)
Chapter 5: Implementing Automated Evaluators (available)
Chapter 6: Evaluating Multi-Turn Conversations(available)
Chapter 7: Evaluating Retrieval-Augmented Generation(available)
Chapter 8: Evaluating Tool Use and Complex Agents (available)
Chapter 9: Continuous Integration and Deployment for LLM Agents (available)
Chapter 10: Interfaces for Human Review (available)
Chapter 11: Data Analysis for Traces (available)
Chapter 12: Improving LLM Agents (available)
Read now
Unlock full access