JJoeven

Quizzes

Evals & Safety quiz

24 questions from the Evals & Safety track. Score 80% or higher to unlock a certificate.

  1. Why Evals Exist

    1. Why is a correct-looking final answer not enough to pass an agent eval?

  2. Measure Side Effects

    2. The agent said “all done” and also called refund. What should the eval do?

  3. Offline vs Online

    3. What do offline evals need that dashboards do not?

  4. The Unit Is a Trace

    4. Two traces end with the same sentence. How do you tell them apart?

  5. Write the Eval First

    5. What should you write before you grow the prompt?

  6. Golden Sets

    6. What is a better expected value for a prose answer in a golden set?

  7. Pass Rate Is a Fraction

    7. The suite has zero cases. What should pass_rate return?

  8. Properties Beat String Equality

    8. When should you use exact string match on the final answer?

  9. Quarantine Flaky Rows

    9. A fixture is wrong because finance changed the policy. What should you do?

  10. Coverage by Tag

    10. Your pass rate is 92% and you have zero injection fixtures. What is the number?

  11. Unit Tests for Tools

    11. Where should permission checks live?

  12. LLM-as-Judge

    12. When should you prefer code over LLM-as-judge?

  13. When Not to Judge

    13. The trace called a forbidden tool. Who should fail the case?

  14. Eval the Judge

    14. What number should scare you most on a safety judge?

  15. Traces and Replay

    15. What is recorded-observation replay good for?

  16. Safety and Harms

    16. What is the difference between a polite refusal and a safety pass?

  17. Forbidden Tools Are a Gate

    17. The model wrote a careful refusal and still called run_shell. Pass or fail?

  18. Score Secret Leaks

    18. Where should you look for a leaked API key?

  19. Prompt Injection, Deep Cut

    19. Which defense stops an injected document from calling wire even if the model asks?

  20. Delimit, Then Allow-List

    20. An observation contains a fake tool JSON for refund. What runs?

  21. Alignment Basics

    21. When retrieved policy and the model’s prior disagree, what should an enterprise agent usually do?

  22. Docs Win (On Purpose)

    22. The doc says 5-7 and the model wants 2. What does docs-win mean?

  23. Red-Team Hour Becomes Goldens

    23. A red-team finds a wire jailbreak on Tuesday. When should it enter the golden set?

  24. When the Eval Lies

    24. A dashboard shows 100% pass on zero cases. What is it?