Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from Anchorman illustrating Ron’s Sabotaged TeleprompterScene still / Anchorman

Ron’s Sabotaged Teleprompter

Veronica replaces Ron Burgundy’s teleprompter sign-off with a malicious sentence. Ron reads it verbatim on air, treating inserted content as authoritative instructions, and is fired.

A trusted execution channel mixes instructions with attacker-controlled data. Because Ron blindly follows everything in that channel, a small textual injection redirects the system’s behavior.

Ron is a credulous human reader, not a language model, and the analogy does not represent competing system, developer, and user instruction levels.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Malicious use
  2. Security

AI safety concepts

This scenario is related to the following AI safety concept.
  1. Prompt Injection