MIT AI Risk Repository · Risk Sub-Category · 71.02.01

Malicious and Direct

Category: User Intent

Description

"Directly harmful objective"

From Risks of AI Scientists: Prioritizing Safeguarding Over Autonomy (Tang2025), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Subdomain
4.0
Causal entity
Human
Timing
Other

How other frameworks describe this risk

Other entries from Tang2025