MIT AI Risk Repository · Risk Category · 17.01.00
Discrimination, Exclusion and Toxicity
Description
"Social harms that arise from the language model producing discriminatory or exclusionary speech"
From Ethical and social risks of harm from language models (Weidinger2021), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Subdomain
- 1.0
- Causal entity
- AI
- Intent
- Unintentional
- Timing
- Post-deployment
How other frameworks describe this risk
Other entries from Weidinger2021
- Social stereotypes and unfair discrmination
- Social stereotypes and unfair discrmination
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Toxic language
- Lower performance for some languages and social groups
- Lower performance for some languages and social groups
- Information Hazards
- Compromising privacy by leaking private infiormation
- Compromising privacy by leaking private infiormation