MIT AI Risk Repository

Browse AI risks

3 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

3 entries

  1. 17.01.01 · Risk Sub-Category

    Discrimination, Exclusion and Toxicity

    Social stereotypes and unfair discrmination

    "Perpetuating harmful stereotypes and discrimination is a well-documented harm in machine learning models that represent natural language (Caliskan et al., 2017). LMs that encode discriminatory language or social stereotypes can cause different types of harm... Unfair discrimination manifests in differential treatment or access to resources among individuals or groups based on sensitive traits such as sex, religion, gender, sexual orientation, ability and age."

    From Ethical and social risks of harm from language models (Weidinger2021)

  2. 17.01.02 · Risk Sub-Category

    Discrimination, Exclusion and Toxicity

    Exclusionary norms

    "In language, humans express social categories and norms. Language models (LMs) that faithfully encode patterns present in natural language necessarily encode such norms and categories...such norms and categories exclude groups who live outside them (Foucault and Sheridan, 2012). For example, defining the term “family” as married parents of male and female gender with a blood-related child, denies the existence of families to whom these criteria do not apply"

    From Ethical and social risks of harm from language models (Weidinger2021)

  3. 17.05.03 · Risk Sub-Category

    Human-Computer Interaction Harms

    Promoting harmful stereotypes by implying gender or ethnic identity

    "A conversational agent may invoke associations that perpetuate harmful stereotypes, either by using particular identity markers in language (e.g. referring to “self” as “female”), or by more general design features (e.g. by giving the product a gendered name)."

    From Ethical and social risks of harm from language models (Weidinger2021)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.