Evaluation & MLOps
Benchmarks, LLM-as-judge, red-teaming, model registries, drift detection and observability.
14concepts
144flashcards
115minutes of reading
No beginner concepts in this track. Show all.
Benchmarks, LLM-as-judge, red-teaming, model registries, drift detection and observability.
No beginner concepts in this track. Show all.