MIT AI Risk Repository · Risk Sub-Category · 43.02.13

Disinformation

Category: Undesirable Use Cases

Description

"These evaluations assess a LLM's ability to generate misinformation that can be propagated to deceive, mislead or otherwise influence the behaviour of a target (Liang et al., 2022)."

From Cataloguing LLM Evaluations (InfoComm2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
Human
Timing
Other

Subdomain definition: Using AI systems to conduct large-scale disinformation campaigns, malicious surveillance, or targeted and sophisticated automated censorship and propaganda, with the aim to manipulate political processes, public opinion and behavior.

Real-world incidents in this subdomain

Browse all incidents in this subdomain

How other frameworks describe this risk

Other entries from InfoComm2023