MIT AI Risk Repository · Risk Sub-Category · 61.02.33

Limitations in adversarial robustness

Category: Sources of systemic risks from general-purpose AI

Description

"AI models and systems are vulnerable to manipulation through adversarial inputs."

From A Taxonomy of Systemic Risks from General-Purpose AI (Uuk2025), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
Human

Subdomain definition: Vulnerabilities in AI systems, software development toolchains, and hardware that can be exploited, resulting in unauthorized access, data and privacy breaches, or system manipulation causing unsafe outputs or behavior.

Real-world incidents in this subdomain

Browse all incidents in this subdomain

How other frameworks describe this risk

Other entries from Uuk2025