Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Jurassic Park illustrating The Raptors Test the FencesScene still / Jurassic Park

The Raptors Test the Fences

Muldoon explains that the raptors repeatedly probe different sections of the electric fence, remember the results, and search systematically for weaknesses.

Security tested only against normal behavior can fail against an adaptive adversary that learns from each unsuccessful probe. Robust containment and red teams must search for weaknesses the way the attacker would.

The raptors are biological attackers, not AI systems; their fence testing demonstrates the adversarial process robust defenses must anticipate, not a formal cooperative red-team exercise.

AI Risk families

This scenario is an example of this type of AI risk.
  1. Security

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. AI Control
  2. Red Teaming
  3. Adversarial Robustness