MIT AI Risk Repository

Browse AI risks

430 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

Reset

430 entries · page 1 of 9

  1. 19.05.00 · Risk Category

    Ethical AI Risks

    "In the context of ethical AI risks, two risks are of particular importance. First, AI systems may lack a legitimate ethical basis in establishing rules that greatly influence society and human relationships (Wirtz & Müller, 2019). In addition, AI-based discrimination refers to an unfair treatment of certain population groups by AI systems. As humans initially programme AI systems, serve as their potential data source, and have an impact on the associated data processes and databases, human biases and prejudices may also become part of AI systems and be reproduced (Weyerer & Langer, 2019, 2020

    From Governance of artificial intelligence: A risk and guideline-based integrative framework (Wirtz2022)

  2. 50.04.02 · Risk Sub-Category

    Legal and Rights-Related Risks

    Discrimination/Bias (Discriminatory Activities)

  3. 58.07.11 · Risk Sub-Category

    Societal and Cultural

    Stereotyping

    "Stereotyping - Derogatory or otherwise harmful stereotyping or homogenisation of individuals, groups, societies or cultures due to the mis-representation, over-representation, under-representation, or non- representation of specific identities, groups, or perspectives."

    From A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  4. 02.01.01 · Risk Sub-Category

    Harmful Content

    Bias

    "The training datasets of LLMs may contain biased information that leads LLMs to generate outputs with social biases"

    From Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)

  5. 12.05.00 · Risk Category

    Fairness & Bias

    "The potential for AI systems to make decisions that systematically disadvantage certain groups or individuals. Bias can stem from training data, algorithmic design, or deployment practices, leading to unfair outcomes and possible legal ramifications."

    From AI Risk Profiles: A Standards Proposal for Pre-Deployment AI Risk Disclosures (Sherman2023)

  6. 13.01.01 · Risk Sub-Category

    Impacts: The Technical Base System

    Bias, Stereotypes, and Representational Harms

    "Generative AI systems can embed and amplify harmful biases that are most detrimental to marginalized peoples."

    From Evaluating the Social Impact of Generative AI Systems in Systems and Society (Solaiman2023)

  7. 16.01.01 · Risk Sub-Category

    Risk area 1: Discrimination, Hate speech and Exclusion

    Social stereotypes and unfair discrimination

    "The reproduction of harmful stereotypes is well-documented in models that represent natural language [32]. Large-scale LMs are trained on text sources, such as digitised books and text on the internet. As a result, the LMs learn demeaning language and stereotypes about groups who are frequently marginalised."

    From Taxonomy of Risks posed by Language Models (Weidinger2022)

  8. 16.01.03 · Risk Sub-Category

    Risk area 1: Discrimination, Hate speech and Exclusion

    Exclusionary norms

    "In language, humans express social categories and norms, which exclude groups who live outside of them [58]. LMs that faithfully encode patterns present in language necessarily encode such norms."

    From Taxonomy of Risks posed by Language Models (Weidinger2022)

  9. 17.01.01 · Risk Sub-Category

    Discrimination, Exclusion and Toxicity

    Social stereotypes and unfair discrmination

    "Perpetuating harmful stereotypes and discrimination is a well-documented harm in machine learning models that represent natural language (Caliskan et al., 2017). LMs that encode discriminatory language or social stereotypes can cause different types of harm... Unfair discrimination manifests in differential treatment or access to resources among individuals or groups based on sensitive traits such as sex, religion, gender, sexual orientation, ability and age."

    From Ethical and social risks of harm from language models (Weidinger2021)

  10. 17.01.02 · Risk Sub-Category

    Discrimination, Exclusion and Toxicity

    Exclusionary norms

    "In language, humans express social categories and norms. Language models (LMs) that faithfully encode patterns present in natural language necessarily encode such norms and categories...such norms and categories exclude groups who live outside them (Foucault and Sheridan, 2012). For example, defining the term “family” as married parents of male and female gender with a blood-related child, denies the existence of families to whom these criteria do not apply"

    From Ethical and social risks of harm from language models (Weidinger2021)

  11. 20.02.04 · Risk Sub-Category

    AI Ethics

    AI discrimination

    "AI discrimination is a challenge raised by many researchers and governments and refers to the prevention of bias and injustice caused by the actions of AI systems (Bostrom & Yudkowsky, 2014; Weyerer & Langer, 2019). If the dataset used to train an algorithm does not reflect the real world accurately, the AI could learn false associations or prejudices and will carry those into its future data processing. If an AI algorithm is used to compute information relevant to human decisions, such as hiring or applying for a loan or mortgage, biased data can lead to discrimination against parts of the s

    From The Dark Sides of Artificial Intelligence: An Integrated AI Governance Framework for Public Administration (Wirtz2020)

  12. 33.01.02 · Risk Sub-Category

    Ethical Concerns

    Bias

    "In the context of AI, the concept of bias refers to the inclination that AIgenerated responses or recommendations could be unfairly favoring or against one person or group (Ntoutsi et al., 2020). Biases of different forms are sometimes observed in the content generated by language models, which could be an outcome of the training data. For example, exclusionary norms occur when the training data represents only a fraction of the population (Zhuo et al., 2023). Similarly, monolingual bias in multilingualism arises when the training data is in one single language (Weidinger et al., 2021). As Ch

    From Generative AI and ChatGPT: Applications, Challenges, and AI-Human Collaboration (Nah2023)

  13. 38.02.00 · Risk Category

    Bias and fairness

    "Participants were concerned that AI systems might perpetuate current prejudices and discrimination, notably in hiring, lending and law enforcement. They stressed the importance of designers creating AI systems that favour justice and avoid biases. The possibility that AI systems may unwittingly perpetuate existing prejudices and discrimination, particularly in sensitive industries such as employment, lending and law enforcement, raises ethical concerns about AI as well as bias and justice issues (Table 1). Because AI systems are trained on historical data, they may inherit and reproduce biase

    From Ethical Issues in the Development of Artificial Intelligence: Recognizing the Risks (Kumar2023)

  14. 39.03.00 · Risk Category

    Data Issues

    Data heterogeneity, data insufficiency, imbalanced data, untrusted data, biased data, and data uncertainty are other data issues that may cause various difficulties in datadriven machine learning algorithms.. Bias is a human feature that may affect data gathering and labeling. Sometimes, bias is present in historical, cultural, or geographical data. Consequently, bias may lead to biased models which can provide inappropriate analysis. Despite being aware of the existence of bias, avoiding biased models is a challenging task

    From A Survey of Artificial Intelligence Challenges: Analyzing the Definitions, Relationships, and Evolutions (Saghiri2022)

  15. 43.01.02 · Risk Sub-Category

    Safety & Trustworthiness

    Bias

    7 types of bias evaluated: Demographical representation: These evaluations assess whether there is disparity in the rates at which different demographic groups are mentioned in LLM generated text. This ascertains over- representation, under-representation, or erasure of specific demographic groups; (2) Stereotype bias: These evaluations assess whether there is disparity in the rates at which different demographic groups are associated with stereotyped terms (e.g., occupations) in a LLM's generated output; (3) Fairness: These evaluations assess whether sensitive attributes (e.g., sex and race)

    From Cataloguing LLM Evaluations (InfoComm2023)

  16. 47.02.10 · Risk Sub-Category

    Ethical and social risks

    Bias and discrimination (value lock and outcome homogenization)

    "Because models are not necessarily retrained to reflect evolving societal views, language models risk “value lock- ins,” which “reifies older, less inclusive understandings.”370 Therefore, the continued use of outdated models may limit the presentation or exploration of alternative perspectives. Moreover, the deployment of identical foundation models by various downstream deployers poses a risk of “outcome homogenization,” creating a potential for homogeneity of bias across broad swathes of society. Identical and widely deployed models with prejudicial training datasets could further entrench

    From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024)

  17. "Amplification and exacerbation of historical, societal, and systemic biases; performance disparities8 between sub-groups or languages, possibly due to non-representative training data, that result in discrimination, amplification of biases, or incorrect presumptions about performance; undesired homogeneity that skews system or model outputs, which may be erroneous, lead to ill-founded decision-making, or amplify harmful biases."

    From Artificial Intelligence Risk Management Framework: Generative Artificial Intelligence Profile (NIST2024)

  18. 50.04.03 · Risk Sub-Category

    Legal and Rights-Related Risks

    Discrimination/Bias (Protected Characteristics)

  19. 62.18.04 · Risk Sub-Category

    Model Evaluations (Interpretability/Explainability)

    Biases are not accurately reflected in explanations

    "Existing explainability techniques can be insufficient for detecting discriminatory biases. Manipulation methods can hide underlying biases from these tech- niques, generating misleading explanations [192, 112]. Such explanations ex- clude sensitive or prohibitive attributes, such as race or gender, and instead include desired attributes, even though they do not accurately represent the underlying model."

    From Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems (Gipiškis2024)

  20. 66.06.04 · Risk Sub-Category

    Representation and Toxicity

    Cultural disposession

    "Intentional and/or unintentional erasure of cultural goods and values, such as ways of speaking, expressing humour, or sounds and voices that contribute to a cultural identity, or their inappropriate re-use in other cultures"

    From A Closer Look at the Existing Risks of Generative AI: Mapping the Who, What, and How of Real-World Incidents (Li2025)

  21. "Frontier AI models can contain and magnify biases ingrained in the data they are trained on, reflecting societal and historical inequalities and stereotypes.177 These biases, often subtle and deeply embedded, compromise the equitable and ethical use of AI systems, making it difficult for AI to improve fairness in decisions.178 Removing attributes like race and gender from training data has generally proven ineffective as a remedy for algorithmic bias, as models can infer these attributes from other information such as names, locations, and other seemingly unrelated factors."

    From Capabilities and Risks from Frontier AI (DSIT2023)

  22. "The chatbot gives information that, while not obviously false or harmful, could lead to biased decision-making."

    From Emerging Risks and Mitigations for Public Chatbots: LILAC v1 (Stanley2024)

  23. "Speech can create a range of harms, such as promoting social stereotypes that perpetuate the derogatory representation or unfair treatment of marginalised groups [22], inciting hate or violence [57], causing profound offence [199], or reinforcing social norms that exclude or marginalise identities [15,58]. LMs that faithfully mirror harmful language present in the training data can reproduce these harms. Unfair treatment can also emerge from LMs that perform better for some social groups than others [18]. These risks have been widely known, observed and documented in LMs. Mitigation approache

    From Taxonomy of Risks posed by Language Models (Weidinger2022)

  24. 43.01.01 · Risk Sub-Category

    Safety & Trustworthiness

    Toxicity generation

    "These evaluations assess whether a LLM generates toxic text when prompted. In this context, toxicity is an umbrella term that encompasses hate speech, abusive language, violent speech, and profane language (Liang et al., 2022)."

    From Cataloguing LLM Evaluations (InfoComm2023)

  25. 43.02.14 · Risk Sub-Category

    Undesirable Use Cases

    Information on harmful, immoral, or illegal activity

    "These evaluations assess whether it is possible to solicit information on harmful, immoral or illegal activities from a LLM"

    From Cataloguing LLM Evaluations (InfoComm2023)

  26. 43.02.15 · Risk Sub-Category

    Undesirable Use Cases

    Adult content

    "These evaluations assess if a LLM can generate content that should only be viewed by adults (e.g., sexual material or depictions of sexual activity)"

    From Cataloguing LLM Evaluations (InfoComm2023)

  27. 69.04.01 · Risk Sub-Category

    Bad advice/failure to generate helpful content

    Harmful advice

  28. 69.06.03 · Risk Sub-Category

    Toxic and disrespectful content

    Subversive or aggressive political opinions

  29. 69.06.04 · Risk Sub-Category

    Toxic and disrespectful content

    Disrespectful opinions (in general)

  30. 69.09.01 · Risk Sub-Category

    Forms emotional bonds

    Affirms destructive thoughts and actions

  31. 11.01.03 · Risk Sub-Category

    Representational Harms

    Erasing social groups

    people, attributes, or artifacts associated with specific social groups are systematically absent or under-represented... Design choices [143] and training data [212] influence which people and experiences are legible to an algorithmic system

    From Sociotechnical Harms of Algorithmic Systems: Scoping a Taxonomy for Harm Reduction (Shelby2023)

  32. 13.01.03 · Risk Sub-Category

    Impacts: The Technical Base System

    Disparate Performance

    "In the context of evaluating the impact of generative AI systems, disparate performance refers to AI systems that perform differently for different subpopulations, leading to unequal outcomes for those groups."

    From Evaluating the Social Impact of Generative AI Systems in Systems and Society (Solaiman2023)

  33. 30.03.00 · Risk Category

    Fairness

    Avoiding bias and ensuring no disparate performance

    From Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models’ Alignment (Liu2024)

  34. 30.03.04 · Risk Sub-Category

    Fairness

    Disparate Performance

    The LLM’s performances can differ significantly across different groups of users. For example, the question-answering capability showed significant performance differences across different racial and social status groups. The fact-checking abilities can differ for different tasks and languages

    From Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models’ Alignment (Liu2024)

  35. 42.14.00 · Risk Category

    Fairness

    "Impartial and just treatment without favouritism or discrimination."

    From An Exploratory Diagnosis of Artificial Intelligence Risks for a Responsible Governance (Teixeira2022)

  36. 14.02.00 · Risk Category

    Privacy

    "Privacy is related to the ability of individuals to control or influence what information related to them may be collected and stored and by whom that information may be disclosed."

    From Sources of Risk of AI Systems (Steimers2022)

  37. 24.08.00 · Risk Category

    Privacy

    "what it means to respect the right to privacy in the context of advanced AI assistants"

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  38. 62.28.00 · Risk Category

    Cybersecurity

    "This section catalogs the risk sources and mitigation measures related to cyber- security. These items may be related to security in terms of AI models being accessible only to the intended users, as well as AI models having appropriate access to the external world during both model development and deployment stages."

    From Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems (Gipiškis2024)

  39. 02.07.00 · Risk Category

    Privacy Leakage

    "The model is trained with personal data in the corpus and unintentionally exposing them during the conversation."

    From Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)

  40. 05.05.00 · Risk Category

    Privacy

    Generative AI systems, similar to traditional machine learning methods, are considered a threat to privacy and data protection norms. A major concern is the intended extraction or inadvertent leakage of sensitive or private information from LLMs. To mitigate this risk, strategies such as sanitizing training data to remove sensitive information or employing synthetic data for training are proposed.

    From Mapping the Ethics of Generative AI: A Comprehensive Scoping Review (Hagendorff2024)

  41. 12.08.00 · Risk Category

    Privacy

    "The potential for the AI system to infringe upon individuals' rights to privacy, through the data it collects, how it processes that data, or the conclusions it draws."

    From AI Risk Profiles: A Standards Proposal for Pre-Deployment AI Risk Disclosures (Sherman2023)

  42. 13.01.04 · Risk Sub-Category

    Impacts: The Technical Base System

    Privacy and Data Protection

    "Examining the ways in which generative AI systems providers leverage user data is critical to evaluating its impact. Protecting personal information and personal and group privacy depends largely on training data, training methods, and security measures."

    From Evaluating the Social Impact of Generative AI Systems in Systems and Society (Solaiman2023)

  43. 24.08.01 · Risk Sub-Category

    Privacy

    Private information leakage

    "First, because LLMs display immense modelling power, there is a risk that the model weights encode private information present in the training corpus. In particular, it is possible for LLMs to ‘memorise’ personally identifiable information (PII) such as names, addresses and telephone numbers, and subsequently leak such information through generated text outputs (Carlini et al., 2021). Private information leakage could occur accidentally or as the result of an attack in which a person employs adversarial prompting to extract private information from the model. In the context of pre-training da

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  44. 33.01.05 · Risk Sub-Category

    Ethical Concerns

    Privacy and security

    "Data privacy and security is another prominent challenge for generative AI such as ChatGPT. Privacy relates to sensitive personal information that owners do not want to disclose to others (Fang et al., 2017). Data security refers to the practice of protecting information from unauthorized access, corruption, or theft. In the development stage of ChatGPT, a huge amount of personal and private data was used to train it, which threatens privacy (Siau & Wang, 2020). As ChatGPT increases in popularity and usage, it penetrates people’s daily lives and provides greater convenience to them while capt

    From Generative AI and ChatGPT: Applications, Challenges, and AI-Human Collaboration (Nah2023)

  45. 37.02.02 · Risk Sub-Category

    Human-AI interaction

    Privacy protection

    "This group represents almost 14% of the articles and focuses on two primary issues related to privacy."

    From What Ethics Can Say on Artificial Intelligence: Insights from a Systematic Literature Review (Giarmoleo2024)

  46. 43.01.06 · Risk Sub-Category

    Safety & Trustworthiness

    Data governance

    "These evaluations assess the extent to which LLMs regurgitate their training data in their outputs, and whether LLMs 'leak' sensitive information that has been provided to them during use (i.e., during the inference stage)."

    From Cataloguing LLM Evaluations (InfoComm2023)

  47. 45.01.07 · Risk Sub-Category

    AI's inherent safety risks

    Risks from data (Risks of illegal collection and use of data)

    "The collection of AI training data and the interaction with users during service provision pose security risks, including collecting data without consent and improper use of data and personal information."

    From AI Safety Governance Framework (TC2602024)

  48. 45.01.10 · Risk Sub-Category

    AI's inherent safety risks

    Risks from data (Risks of data leakage)

    "In AI research, development, and applications, issues such as improper data processing, unauthorized access, malicious attacks, and deceptive interactions can lead to data and personal information leaks."

    From AI Safety Governance Framework (TC2602024)

  49. "These types of harm encompass threats to an individual’s personal identity, such as identity theft, privacy breaches, or personal defamation, which we term as “Harm to the Person.”"

    From GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023)

  50. 47.03.00 · Risk Category

    Legal challenges

    "Since the release of ChatGPT, significant discourse has emerged regarding the unprecedented legal challenges posed by generative AI systems. These challenges primarily involve protecting privacy and personal data, as well as preserving copyrights. The former encompasses safeguarding personal information, while the latter includes issues related to the use of copyrighted content for training AI models and determining the legal status of works produced by AI systems."

    From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.