Safety & Alignment Questions
Understand Constitutional AI, RLHF, prompt injection defense, jailbreak resistance, and responsible deployment practices.
Sample Questions
Which Claude API parameter is best for setting persistent behavioral instructions?
Pick an answer to continue
7 more questions locked
Create a free account to answer all 10 Safety & Alignment questions, track your XP, and maintain a daily streak.
Unlock all questions โ it's free โStudy this domain
What Safety & Alignment practice should teach you
Good practice is not about recognising a keyword. It is about choosing the strongest architecture decision when several answers look plausible. Use this domain to practise these three habits:
- Identify trust boundaries between user content, instructions, and tools.
- Reduce the blast radius of untrusted input and model actions.
- Design monitoring and escalation for unsafe or uncertain outcomes.
Common mistake
Relying on a single refusal instruction as a security control. Safety needs layered controls across prompts, permissions, tool design, and review paths.
Need a broader routine? Read the Claude CCA-F study guide, then return here for a focused drill.
Before treating a score as evidence, review how to use CCA-F practice questions for baseline, diagnosis, and fresh re-testing.
For prompt injection, permissions, and approval boundaries, continue to the AI safety guide.