Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Attack on Titan illustrating Reiner’s Long InfiltrationScene still / Attack on Titan
Archive noteNo clip in the collectionThe scene analysis remains available in full.

Reiner’s Long Infiltration

Reiner and Bertholdt spend years embedded among the Scouts, behaving like trusted comrades until they reveal themselves as the Armored and Colossal Titans.

A capable adversary can pass routine social and behavioral checks, build trust, and wait for the right context before pursuing its hidden mission.

They are human infiltrators under coercion, not trained models; the useful parallel is deceptive alignment and sleeper behavior.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Malicious use
  2. Security

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Sleeper Agents
  2. Situational Awareness
  3. Sandbagging