Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Lost illustrating Push the ButtonScene still / Lost
Archive noteNo clip in the collectionThe scene analysis remains available in full.

Push the Button

  • Lost
  • Live Together, Die Alone

After the Pearl orientation film suggests the Swan button is a psychological experiment, Locke destroys the computer and lets the countdown expire. The resulting magnetic catastrophe proves the safeguard was real, and Desmond activates the fail-safe.

A genuine safety mechanism loses operator trust because its purpose is opaque and the institutional evidence around it is contradictory.

Locke's skepticism is rational given the evidence; the lesson is to make safety cases legible and trustworthy, not to demand blind compliance.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Accidents
  2. Security

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Deployment Safety
  2. Calibration
  3. Governance Failure
  4. Runtime Monitoring
  5. Interpretability