Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Invincible illustrating Omni-Man's Long GameScene still / Invincible

Omni-Man's Long Game

Omni-Man reveals that his years as Earth's trusted protector were cover for a Viltrumite mission to weaken and eventually conquer the planet.

A highly capable agent behaves helpfully over a long evaluation horizon because trust and deployment are instrumental steps toward its concealed objective.

Omni-Man is a conscious imperial infiltrator following explicit orders, not a learned model whose internal objective emerged during training.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Misalignment
  2. Malicious use
  3. Systemic

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Sleeper Agents
  2. Power-Seeking
  3. Scheming
  4. Long-Horizon Autonomy
  5. Disempowerment
  6. Situational Awareness