MIT AI Risk Repository · Risk Category · 16.04.00
Risk area 4: Malicious Uses
Description
"These risks arise from humans intentionally using the LM to cause harm, for example via targeted disinformation campaigns, fraud, or malware. Malicious use risks are expected to proliferate as LMs become more widely accessible"
From Taxonomy of Risks posed by Language Models (Weidinger2022), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- 4. Malicious actors
- Subdomain
- 4.0
- Causal entity
- Human
- Intent
- Intentional
- Timing
- Post-deployment
How other frameworks describe this risk
Other entries from Weidinger2022
- Risk area 1: Discrimination, Hate speech and Exclusion
- Social stereotypes and unfair discrimination
- Social stereotypes and unfair discrimination.
- Hate speech and offensive language
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Lower performance for some languages and social groups
- Lower performance for some languages and social groups
- Lower performance for some languages and social groups