Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Parks and Recreation illustrating T-Dazzle

T-Dazzle

Leslie and Tom rebrand fluoride as “T-Dazzle,” turning an unpopular public-health policy into a 72-percent polling winner without changing the substance.

An interface optimized around biases can manufacture apparent preference without producing better-informed consent.

The tactic supports a beneficial policy, which makes the manipulation-versus-outcome tradeoff especially ambiguous.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Malicious use
  2. Systemic

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Value Alignment
  2. AI Disinformation
  3. Sycophancy