Cultural Alignment

Exploring real AI risks through the lens of pop culture

Still from The Wire illustrating Xerox Lie Detector

Xerox Lie Detector

Detectives tape a suspect’s hands to a photocopier preloaded with “TRUE” and “FALSE” pages and convince him the office machine can read lies; he confesses.

Users may grant epistemic authority to a system they do not understand, making the appearance of automation as influential as real capability.

The machine is deliberately fake and the police exploit the suspect’s misunderstanding.

AI Risk families

This scenario is an example of the following types of AI risk.
  1. Malicious use
  2. Systemic

AI safety concepts

This scenario is related to the following AI safety concepts.
  1. Automation Bias
  2. Synthetic Fraud
  3. Interpretability