Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from The Boys illustrating Ryan’s Unsafe Demo

Ryan’s Unsafe Demo

Vought scripts a harmless first “save,” but Homelander changes the instructions mid-demonstration; Ryan follows the command and splatters his stunt coach against a building.

A staged evaluation becomes live deployment without a calibrated capability estimate, stable protocol, or safety margin.

Ryan is a superpowered child under coercive human direction, not an autonomous AI system.

AI Risk families

This scenario is an example of this type of AI risk.
  1. Accidents

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Deployment Safety
  2. Capability Evals
  3. Calibration
  4. Side Effects