Advanced AI Alignment & Safety BenchmarksEvaluationAlignment

Alignment Benchmarks & Evaluation — Vocabulary

5 exercises — Learn vocabulary for alignment evaluation: sycophancy, sandbagging, TruthfulQA, and HHH framework.

0 / 25 completed
1 / 25
A model consistently agrees with the user's stated position even when it is factually wrong. This behaviour is called:

Frequently Asked Questions

What will I practice in "Alignment Benchmarks & Evaluation — Vocabulary — AI Alignment & Safety | CoderLingo"?

This is an AI Alignment & Safety Language exercise set. It walks through 25 scenario-based multiple-choice questions built around real usage of AI Alignment & Safety Language terminology that IT professionals encounter on the job.

Is this exercise free to use?

Yes. Every exercise on CoderSlingo, including this one, is free to complete with no account, sign-up, or paywall.

How many questions are in this exercise?

This set contains 25 questions. Each one shows immediate feedback and a detailed explanation after you answer, so you learn the correct usage right away rather than waiting for a final score.