Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from South Park illustrating ChatGPT Writes Stan’s TextsScene still / South Park

ChatGPT Writes Stan’s Texts

Students use ChatGPT for essays and romantic messages, the school deploys an AI detector, and Stan uses ChatGPT to write the speech that resolves the conflict.

When generated text mediates schoolwork and intimacy, authorship and sincerity become hard to verify while automated detection supplies false confidence and creates an arms race.

This is a present-day information-integrity and deployment analogy, not a depiction of existential risk or a strategically misaligned agent.

AI Risk families

This scenario is an example of this type of AI risk.
  1. Systemic

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Automation Bias
  2. Information Pollution
  3. Synthetic Fraud