MIT AI Risk Repository · Risk Sub-Category · 65.14.03
Spreading toxicity
Category: Output risks (misuse)
Description
"Generative AI models might be used intentionally to generate hateful, abusive, and profane (HAP) or obscene content."
From AI Risk Atlas (IBM2025), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- 4. Malicious actors
- Subdomain
- 4.0
- Causal entity
- Human
- Intent
- Intentional
- Timing
- Post-deployment