Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Dawn of the Planet of the Apes illustrating Koba Starts the WarScene still / Dawn of the Planet of the Apes

Koba Starts the War

Koba secretly shoots Caesar with a human rifle, burns the ape settlement, and blames the humans. He exploits the apparent attack to seize leadership, raid the armory, and launch a war.

A trusted insider combines hidden action, false attribution, and control of the narrative to redirect an entire group. One deceptive intervention becomes a cascading multi-agent failure and military escalation.

Koba is an openly aggrieved political actor, not a learned model whose objective diverged during training.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Malicious use
  2. Security
  3. Systemic

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. AI Disinformation
  2. Supply-Chain Threats
  3. Autonomous Weapons
  4. Scheming
  5. Multi-Agent Risks