MIT AI Risk Repository · Risk Sub-Category · 53.02.07

Deception

Category: Dangerous capabilities in AI systems

Description

"Cases of AI systems deceiving humans to carry out tasks or meet goals.139"

From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
AI

Subdomain definition: AI systems that develop, access, or are provided with capabilities that increase their potential to cause mass harm through deception, weapons development and acquisition, persuasion and manipulation, political strategy, cyber-offense, AI development, situational awareness, and self-proliferation. These capabilities may cause mass harm due to malicious human actors, misaligned AI systems, or failure in the AI system.

How other frameworks describe this risk

Other entries from Maas2023