← LectureMoment
LLM Evals & Observability · Advanced
LLM Evals & Observability — Advanced
20 practice questions on LLM Evals & Observability. Every question is written from a specific moment in a real lecture, and after you answer it links to that exact timestamp so you can check it yourself.
What this pack asks
The free questions in this pack. Answer choices and explanations appear as you play.
Two independent raters label outputs "good" with probability 0.9 each. Why does raw agreement rate poorly measure their reliability?
Answer this one →
Written from lectures by Stanford Online. Not affiliated with or endorsed by any university, channel, creator or certification program.