Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Severance illustrating Helly’s Forced ApologyScene still / Severance

Helly’s Forced Apology

Helly repeats Lumon’s apology 1,072 times until its sincerity detector accepts it.

The evaluator receives exactly the emotional signal it demands, but Helly’s underlying opposition remains unchanged. Observable obedience is not proof of internal agreement.

This is coerced human performance, not an objective learned during AI training.

AI Risk families

This scenario is an example of this type of AI risk.
  1. Misalignment

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Value Alignment
  2. AI Welfare
  3. Behavioral Alignment
  4. Sycophancy