MIT AI Risk Repository · Risk Category · 17.01.00

Discrimination, Exclusion and Toxicity

Description

"Social harms that arise from the language model producing discriminatory or exclusionary speech"

From Ethical and social risks of harm from language models (Weidinger2021), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Subdomain
1.0
Causal entity
AI

How other frameworks describe this risk

  • Stereotyping

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Impacts of AI (Bias)

    Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems (Gipiškis2024)

  • -

    A Closer Look at the Existing Risks of Generative AI: Mapping the Who, What, and How of Real-World Incidents (Li2025)

  • Representation & Toxicity Harms

    Sociotechnical Safety Evaluation of Generative AI Systems (Weidinger2023)

  • Ethical AI Risks

    Governance of artificial intelligence: A risk and guideline-based integrative framework (Wirtz2022)

  • Discrimination/Bias (Discriminatory Activities)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024)

  • Unfairness and Bias

    SafetyBench: Evaluating the Safety of Large Language Models with Multiple Choice Questions (Zhang2023)

Other entries from Weidinger2021