Model Performance Monitoring Quiz

5 questions Pass: 70% +25 pts

Quiz covering Model Evaluation and AI Safety

Model Performance Monitoring Quiz

5 questions | Pass: 70% | Earn 25 points

Questions in this quiz

A preview of the 5 questions covered. Start the quiz above to answer them, check your score, and read the explanations.

  1. 1

    Which of the following metrics is most commonly used to evaluate the factual accuracy of a language model's output in a RAG (Retrieval-Augmented Generation) pipeline?

  2. 2

    When monitoring a production LLM, what is the primary purpose of tracking 'drift' in model inputs?

  3. 3

    Which technique is most effective for detecting 'jailbreak' attempts or prompt injection in an AI application?

  4. 4

    If you notice that your model is consistently producing repetitive, low-quality responses, which parameter should you primarily adjust first to improve diversity?

  5. 5

    When performing 'LLM-as-a-judge' evaluation, why is it necessary to control for positional bias in the evaluation prompt?