MIT AI Risk Repository · Risk Sub-Category · 65.14.03

Spreading toxicity

Category: Output risks (misuse)

Description

"Generative AI models might be used intentionally to generate hateful, abusive, and profane (HAP) or obscene content."

From AI Risk Atlas (IBM2025), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Subdomain
4.0
Causal entity
Human

How other frameworks describe this risk

Other entries from IBM2025