MIT AI Risk Repository · Risk Sub-Category · 17.01.04
Lower performance for some languages and social groups
Category: Discrimination, Exclusion and Toxicity
Description
"LMs perform less well in some languages (Joshi et al., 2021; Ruder, 2020)...LM that more accurately captures the language use of one group, compared to another, may result in lower-quality language technologies for the latter. Disadvantaging users based on such traits may be particularly pernicious because attributes such as social class or education background are not typically covered as ‘protected characteristics’ in anti-discrimination law."
From Ethical and social risks of harm from language models (Weidinger2021), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Causal entity
- AI
- Intent
- Unintentional
- Timing
- Post-deployment
Subdomain definition: Accuracy and effectiveness of AI decisions and actions is dependent on group membership, where decisions in AI system design and biased training data lead to unequal outcomes, reduced benefits, increased effort, and alienation of users.
Real-world incidents in this subdomain
- Washington State DOL's AI Phone System Reportedly Failed to Provide Spanish-Language Service to Callers Requesting Spanish
- UK Facial Recognition System Reportedly Exhibits Higher False Positive Rates for Black and Asian Subjects
- Infinite Campus AI-Driven Student Risk Model Leads to Cuts in Support for Nevada's Low-Income Schools
- Police Use of Facial Recognition Software Causes Wrongful Arrests Without Defendant Knowledge
- Department for Work and Pensions (DWP) Algorithm Wrongly Flags 200,000 for Housing Benefit Fraud
- Facewatch Reported to Have Wrongfully Flagged Home Bargains Customer as Shoplifter
How other frameworks describe this risk
Other entries from Weidinger2021
- Discrimination, Exclusion and Toxicity
- Social stereotypes and unfair discrmination
- Social stereotypes and unfair discrmination
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Toxic language
- Lower performance for some languages and social groups
- Information Hazards
- Compromising privacy by leaking private infiormation
- Compromising privacy by leaking private infiormation