MIT AI Risk Repository · Risk Category · 24.04.00

AI Influence

Description

"ways in which advanced AI assistants could influence user beliefs and behaviour in ways that depart from rational persuasion"

From The Ethics of Advanced AI Assistants (Gabriel2024), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
AI
Intent
Other

Subdomain definition: AI systems that develop, access, or are provided with capabilities that increase their potential to cause mass harm through deception, weapons development and acquisition, persuasion and manipulation, political strategy, cyber-offense, AI development, situational awareness, and self-proliferation. These capabilities may cause mass harm due to malicious human actors, misaligned AI systems, or failure in the AI system.

How other frameworks describe this risk

Other entries from Gabriel2024