๐Ÿ›ก๏ธ

Safety & Alignment Questions

Understand Constitutional AI, RLHF, prompt injection defense, jailbreak resistance, and responsible deployment practices.

10 questionsยท3 easy ยท 4 medium ยท 3 hard

Sample Questions

Safetyeasy1 / 3

Which Claude API parameter is best for setting persistent behavioral instructions?

Pick an answer to continue

๐Ÿ”’

7 more questions locked

Create a free account to answer all 10 Safety & Alignment questions, track your XP, and maintain a daily streak.

Unlock all questions โ€” it's free โ†’

Study this domain

What Safety & Alignment practice should teach you

Good practice is not about recognising a keyword. It is about choosing the strongest architecture decision when several answers look plausible. Use this domain to practise these three habits:

  • Identify trust boundaries between user content, instructions, and tools.
  • Reduce the blast radius of untrusted input and model actions.
  • Design monitoring and escalation for unsafe or uncertain outcomes.

Common mistake

Relying on a single refusal instruction as a security control. Safety needs layered controls across prompts, permissions, tool design, and review paths.

Need a broader routine? Read the Claude CCA-F study guide, then return here for a focused drill.

Before treating a score as evidence, review how to use CCA-F practice questions for baseline, diagnosis, and fresh re-testing.

For prompt injection, permissions, and approval boundaries, continue to the AI safety guide.

Explore other domains