Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from The Wire illustrating Hamsterdam

Hamsterdam

Colvin concentrates the drug trade into unofficial free zones, sharply improving visible neighborhood crime statistics while concentrating severe harm inside the zones.

A target can improve while omitted values and distributional effects make the outcome morally ambiguous or unacceptable.

Hamsterdam also produces real benefits, so it is an intentionally complex policy tradeoff rather than a clean optimization failure.

AI Risk families

This scenario is an example of this type of AI risk.
  1. Systemic

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Goodhart’s Law
  2. Contestability
  3. Side Effects
  4. Governance Failure