MIT AI Risk Repository · Risk Sub-Category · 53.04.03
Impacts on “epistemic security” and the information environment
Category: Indirect AI contributions to existential risks
Description
-
From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- 3. Misinformation
- Causal entity
- AI
- Intent
- Unintentional
- Timing
- Post-deployment
Subdomain definition: Highly personalized AI-generated misinformation creating “filter bubbles” where individuals only see what matches their existing beliefs, undermining shared reality, weakening social cohesion and political processes.
Real-world incidents in this subdomain
- Google Books Appears to Be Indexing Works Written by AI
- Uptick in Low-Quality AI-Produced Content Degraded Publishers' Submission Management
- Korean Politician Employed Deepfake as Campaign Representative
- Facebook Political Ad Delivery Algorithms Inferred Users' Political Alignment, Inhibiting Political Campaigns' Reach
How other frameworks describe this risk
- Information degradation
- Radicalisation
- Institutional trust loss
- Worsened epistemic processes for society
- Reduced decision-making capacity as a result of decreased trust in information
- Widespread use of persuasive tools contributes to splintered epistemic communities
- AI contributes to increased online polarisation
- Degradation of the information environment
Other entries from Maas2023
- Alignment failures in existing ML systems
- Faulty reward functions in the wild
- Specification gaming
- Reward model overoptimization
- Instrumental convergence
- Goal misgeneralization
- Inner misalignment
- Language model misalignment
- Harms from increasingly agentic algorithmic systems
- Dangerous capabilities in AI systems
- Situational awareness
- Acquisition of a goal to harm society