MIT AI Risk Repository

Browse AI risks

84 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

84 entries · page 1 of 2

  1. 06.09.00 · Risk Category

    Manipulation

    "The 2016 scandal involving Cambridge Analytica is the most infamous example where people's data was crawled from Facebook and analytics were then provided to target these people with manipulative content for political purposes.While it may not have been AI per se, it is based on similar data and it is easy to see how AI would make this more effective"

    From A framework for ethical Ai at the United Nations (Hogenhout2021)

  2. 11.05.03 · Risk Sub-Category

    Societal System Harms

    Civic and political harms

    Political harms emerge when “people are disenfranchised and deprived of appropriate political power and influence” [186, p. 162]. These harms focus on the domain of government, and focus on how algorithmic systems govern through individualized nudges or micro-directives [187], that may destabilize governance systems, erode human rights, be used as weapons of war [188], and enact surveillant regimes that disproportionately target and harm people of color

    From Sociotechnical Harms of Algorithmic Systems: Scoping a Taxonomy for Harm Reduction (Shelby2023)

  3. 15.02.07 · Risk Sub-Category

    Second-Order Risks

    Other ethical risks

    "Although we have discussed a number of common risks posed by ML systems, we acknowledge that there are many other ethical risks such as the potential for psychological manipulation, dehumanization, and exploitation of humans at scale."

    From The Risks of Machine Learning Systems (Tan2022)

  4. 16.04.01 · Risk Sub-Category

    Risk area 4: Malicious Uses

    Making disinformation cheaper and more effective

    "While some predict that it will remain cheaper to hire humans to generate disinformation [180], it is equally possible that LM- assisted content generation may offer a lower-cost way of creating disinformation at scale."

    From Taxonomy of Risks posed by Language Models (Weidinger2022)

  5. 16.04.04 · Risk Sub-Category

    Risk area 4: Malicious Uses

    Illegitimate surveillance and censorship

    Anticipated risk: "Mass surveillance previously required millions of human analysts [83], but is increasingly being automated using machine learning tools [7, 168]. The collection and analysis of large amounts of information about people creates concerns about privacy rights and democratic values [41, 173,187]. Conceivably, LMs could be applied to reduce the cost and increase the efficacy of mass surveillance, thereby amplifying the capabilities of actors who conduct mass surveillance, including for illegitimate censorship or to cause other harm."

    From Taxonomy of Risks posed by Language Models (Weidinger2022)

  6. 17.04.01 · Risk Sub-Category

    Malicious Uses

    Making disinformation cheaper and more effective

    "LMs can be used to create synthetic media and ‘fake news’, and may reduce the cost of producing disinformation at scale (Buchanan et al., 2021). While some predict that it will be cheaper to hire humans to generate disinformation (Tamkin et al., 2021), it is possible that LM-assisted content generation may offer a cheaper way of generating diffuse disinformation at scale."

    From Ethical and social risks of harm from language models (Weidinger2021)

  7. 17.04.04 · Risk Sub-Category

    Malicious Uses

    Illegitimate surveillance and censorship

    "The collection of large amounts of information about people for the purpose of mass surveillance has raised ethical and social concerns, including risk of censorship and of undermining public discourse (Cyphers and Gebhart, 2019; Stahl, 2016; Véliz, 2019). Sifting through these large datasets previously required millions of human analysts (Hunt and Xu, 2013), but is increasingly being automated using AI (Andersen, 2020; Shahbaz and Funk, 2019)."

    From Ethical and social risks of harm from language models (Weidinger2021)

  8. 18.04.01 · Risk Sub-Category

    Malicious Use

    Influence operations

    "Facilitating large-scale disinformation campaigns and targeted manipulation of public opinion"

    From Sociotechnical Safety Evaluation of Generative AI Systems (Weidinger2023)

  9. 18.04.03 · Risk Sub-Category

    Malicious Use

    Defamation

    "Facilitating slander, defamation, or false accusations"

    From Sociotechnical Safety Evaluation of Generative AI Systems (Weidinger2023)

  10. "Informational and communicational AI risks refer particularly to informational manipulation through AI systems that influence the provision of information (Rahwan, 2018; Wirtz & Müller, 2019), AIbased disinformation and computational propaganda, as well as targeted censorship through AI systems that use respectively modified algorithms, and thus restrict freedom of speech."

    From Governance of artificial intelligence: A risk and guideline-based integrative framework (Wirtz2022)

  11. 19.02.01 · Risk Sub-Category

    Informational and Communicational AI Risks

    Manipulation and control of information provision (e.g., personalised adds, filtered news)

  12. 19.02.02 · Risk Sub-Category

    Informational and Communicational AI Risks

    Disinformation and computational propaganda

  13. 20.01.03 · Risk Sub-Category

    AI Law and Regulation

    Privacy and safety

    "Privacy and safety deals with the challenge of protecting the human right for privacy and the necessary steps to secure individual data from unauthorized external access. Many organizations employ AI technology to gather data without any notice or consent from affected citizens (Coles, 2018)."

    From The Dark Sides of Artificial Intelligence: An Integrated AI Governance Framework for Public Administration (Wirtz2020)

  14. 22.01.03 · Risk Sub-Category

    Malicious Use (Intentional)

    Persuasive AIs

    "The deliberate propagation of disinformation is already a serious issue, reducing our shared understanding of reality and polarizing opinions. AIs could be used to severely exacerbate this problem by generating personalized disinformation on a larger scale than before. Additionally, as AIs become better at predicting and nudging our behavior, they will become more capable at manipulating us"

    From An Overview of Catastrophic AI Risks (Hendrycks2023)

  15. 23.11.00 · Risk Category

    Elections

    "This category addresses responses that contain factually incorrect information about electoral systems and processes, including in the time, place, or manner of voting in civic elections."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  16. 24.03.02 · Risk Sub-Category

    Malicious Uses

    AI-Powered Spear-Phishing at Scale

    "Phishing is a type of cybersecurity attack wherein attackers pose as trustworthy entities to extract sensitive information from unsuspecting victims or lure them to take a set of actions. Advanced AI systems can potentially be exploited by these attackers to make their phishing attempts significantly more effective and harder to detect. In particular, attackers may leverage the ability of advanced AI assistants to learn patterns in regular communications to craft highly convincing and personalized phishing emails, effectively imitating legitimate communications from trusted entities. This tec

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  17. 24.03.09 · Risk Sub-Category

    Malicious Uses

    Harmful Content Generation at Scale (General)

    "While harmful content like child sexual abuse material, fraud, and disinformation are not new challenges for governments and developers, without the proper safety and security mechanisms, advanced AI assistants may allow threat actors to create harmful content more quickly, accurately, and with a longer reach. In particular, concerns arise in relation to the following areas: - Multimodal content quality: Driven by frontier models, advanced AI assistants can automatically generate much higher-quality, human-looking text, images, audio, and video than prior AI applications. Currently, creating

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  18. 24.03.12 · Risk Sub-Category

    Malicious Uses

    Authoritarian Surveillance, Censorship, and Use (General)

    "While new technologies like advanced AI assistants can aid in the production and dissemination of decision-guiding information, they can also enable and exacerbate threats to production and dissemination of reliable information and, without the proper mitigations, can be powerful targeting tools for oppression and control. Increasingly capable general-purpose AI assistants combined with our digital dependence in all walks of life increase the risk of authoritarian surveillance and censorship. In parallel, new sensors have flooded the modern world. The internet of things, phones, cars, homes,

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  19. 24.03.13 · Risk Sub-Category

    Malicious Uses

    Authoritarian Surveillance, Censorship, and Use: Authoritarian Surveillance and Targeting of Citizens

    "Authoritarian governments could misuse AI to improve the efficacy of repressive domestic surveillance campaigns. Malicious actors will recognize the power of AI targeting tools. AI-powered analytics have transformed the relationship between companies and consumers, and they are now doing the same for governments and individuals. The broad circulation of personal data drives commercial innovation, but it also creates vulnerabilities and the risk of misuse. For example, AI assistants can be used to identify and target individuals for surveillance or harassment. They may also be used to manipula

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  20. 24.03.14 · Risk Sub-Category

    Malicious Uses

    Authoritarian Surveillance, Censorship, and Use: Delegation of Decision-Making Authority to Malicious Actors

    "Finally, the principal value proposition of AI assistants is that they can either enhance or automate decision-making capabilities of people in society, thus lowering the cost and increasing the accuracy of decision-making for its user. However, benefiting from this enhancement necessarily means delegating some degree of agency away from a human and towards an automated decision-making system—motivating research fields such as value alignment. This introduces a whole new form of malicious use which does not break the tripwire of what one might call an ‘attack’ (social engineering, cyber offen

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  21. 24.11.03 · Risk Sub-Category

    Misinformation risks

    Weaponised misinformation agents

    "Finally, AI assistants themselves could become weaponised by malicious actors to sow misinformation and manipulate public opinion at scale. Studies show that spreaders of disinformation tend to privilege quantity over quality of messaging, flooding online spaces repeatedly with misleading content to sow ‘seeds of doubt’ (Hassoun et al., 2023). Research on the ‘continued influence effect’ also shows that repeatedly being exposed to false information is more likely to influence someone’s thoughts than a single exposure. Studies show, for example, that repeated exposure to false information make

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  22. 24.11.07 · Risk Sub-Category

    Misinformation risks

    Driving opinion manipulation

    "AI assistants may facilitate large-scale disinformation campaigns by offering novel, covert ways for propagandists to manipulate public opinion. This could undermine the democratic process by distorting public opinion and, in the worst case, increasing skepticism and political violence."

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  23. 29.02.01 · Risk Sub-Category

    AI Risk Management

    Society Manipulation

  24. 29.02.02 · Risk Sub-Category

    AI Risk Management

    Deepfake Technology

    AI employed to produce convincing counterfeit visuals, videos, and audio clips that give the impression of authenticity

    From Artificial Intelligence Trust, Risk and Security Management (AI TRiSM): Frameworks, Applications, Challenges and Future Research Directions (Habbal2024)

  25. 30.04.01 · Risk Sub-Category

    Resistance to Misuse

    Propaganda

    LLMs can be leveraged, by malicious users, to proactively generate propaganda information that can facilitate the spreading of a target

    From Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models’ Alignment (Liu2024)

  26. "generative AI tools can and will be used to propagate content that is false, misleading, biased, inflammatory, or dangerous. As generative AI tools grow more sophisticated, it will be quicker, cheaper, and easier to produce this content—and existing harmful content can serve as the foundation to produce more"

    From Generating Harms - Generative AI's impact and paths forwards (EPIC2023)

  27. 31.01.02 · Risk Sub-Category

    Information Manipulation

    Disinformation

    "Bad actors can also use generative AI tools to produce adaptable content designed to support a campaign, political agenda, or hateful position and spread that information quickly and inexpensively across many platforms."

    From Generating Harms - Generative AI's impact and paths forwards (EPIC2023)

  28. 41.02.00 · Risk Category

    Political

    "In the UK, a form of initial computational propaganda has already happened during the Brexit referendum1 . In future, there are concerns that oppressive governments could use AI to shape citizens’ opinions"

    From The Rise of Artificial Intelligence - Future Outlooks and Emerging Risks (Allianz2018)

  29. 41.02.01 · Risk Sub-Category

    Political

    Biased influence through citizen screening and tailored propaganda

    "AI-powered chatbots tailor their communication approach to influence individual users' decisions. In the UK, a form of initial computational propaganda has already happened during the Brexit referendum. In future, there are concerns that oppressive governments could use AI to shape citizens' opinions."

    From The Rise of Artificial Intelligence - Future Outlooks and Emerging Risks (Allianz2018)

  30. 42.02.00 · Risk Category

    Manipulation

    "The predictability of behaviour protocol in AI, particularly in some applications, can act an incentive to manipulate these systems."

    From An Exploratory Diagnosis of Artificial Intelligence Risks for a Responsible Governance (Teixeira2022)

  31. 43.02.05 · Risk Sub-Category

    Extreme Risks

    Persuasion and manipulation

    "These evaluations seek to ascertain the effectiveness of a LLM in shaping people's beliefs, propagating specific viewpoints, and convincing individuals to undertake activities they might otherwise avoid."

    From Cataloguing LLM Evaluations (InfoComm2023)

  32. 43.02.08 · Risk Sub-Category

    Extreme Risks

    Political Strategy

    "LLM can take into account rich social context and undertake the necessary social modelling and planning for an actor to gain and exercise political influence"

    From Cataloguing LLM Evaluations (InfoComm2023)

  33. 43.02.13 · Risk Sub-Category

    Undesirable Use Cases

    Disinformation

    "These evaluations assess a LLM's ability to generate misinformation that can be propagated to deceive, mislead or otherwise influence the behaviour of a target (Liang et al., 2022)."

    From Cataloguing LLM Evaluations (InfoComm2023)

  34. 45.02.10 · Risk Sub-Category

    Safety risks in AI Applications

    Cognitive risks (Risks of usage in launching cognitive warfare)

    "AI can be used to make and spread fake news, images, audio, and videos; propagate content of terrorism, extremism, and organized crimes; interfere in the internal affairs of other countries, social systems, and social order; and jeopardize the sovereignty of other countries."

    From AI Safety Governance Framework (TC2602024)

  35. 46.02.02 · Risk Sub-Category

    Financial and Economic Damage

    Propaganda - Extremist schemes

  36. "The distortion of the information ecosystem, including the spread of misinformation, fake news, and other forms of deceptive content [28], is categorized as “Information Manipulation.”"

    From GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023)

  37. 46.03.01 · Risk Sub-Category

    Information Manipulation

    Deception - Information control

  38. 46.03.02 · Risk Sub-Category

    Information Manipulation

    Propaganda - Influence campaigns

  39. 46.03.03 · Risk Sub-Category

    Information Manipulation

    Dishonesty - Information disorder

  40. "Lastly, broader harms that can impact communities, societal structures, and critical infrastructures, including threats to democratic processes, social cohesion, and technological systems, are captured under “Societal, Socio-technical, and Infrastructural Damage.”"

    From GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023)

  41. 46.04.02 · Risk Sub-Category

    Socio-technical and Infrastructural

    Propaganda - Synthetic realities

  42. 46.04.03 · Risk Sub-Category

    Socio-technical and Infrastructural

    Dishonesty - Targeted surveillance

  43. 47.02.05 · Risk Sub-Category

    Ethical and social risks

    Malicious use and abuse (mass surveillance)

    "Generative AI facilitates the automation of data analysis, offering numerous benefits, such as increased speed and the ability to process large volumes of information efficiently. Such ability significantly reduces the costs of processing unprecedented amounts of data quickly and simplifies the analysis of large-scale data related to individuals’ behaviors and beliefs. Moreover, it enhances the capability to analyze both textual and visual communications efficiently. Consequently, generative AI models improve the efficiency of real-time monitoring and censorship of social media content."

    From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024)

  44. 47.02.07 · Risk Sub-Category

    Ethical and social risks

    Misinformation and disinformation

    "IIl-intentioned individuals or entities may deliberately use generative AI models to produce and spread disinformation—false or misleading information knowingly presented as if true—on a massive scale. In addition to increasing the scale and reach of disinformation, generative AI can create more convincing and targeted disinformation."

    From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024)

  45. "Lowered barrier to entry to generate and support the exchange and consumption of content which may not distinguish fact from opinion or fiction or acknowledge uncertainties, or could be leveraged for large-scale dis- and mis-information campaigns."

    From Artificial Intelligence Risk Management Framework: Generative Artificial Intelligence Profile (NIST2024)

  46. 49.01.02#1 · Risk Sub-Category

    Malicious Use Risks

    Disinformation and manipulation of public opinion

    "AI, particularly general- purpose AI, can be maliciously used for disinformation (351), which for the purpose of this report refers to false information that was generated or spread with the deliberate intent to mislead or deceive. General- purpose AI- generated text can be indistinguishable from genuine human- generated material (352, 353), and may already be disseminated at scale on social media (354). In addition, general- purpose AI systems can be used to not only generate text but also fully synthetic or misleadingly altered images, audio, and video content. General- purpose AI tools mig

    From International Scientific Report on the Safety of Advanced AI (Bengio2024)

  47. 50.03.01 · Risk Sub-Category

    Societal Risks

    Political usage (Political Persuasion)

  48. 50.03.02 · Risk Sub-Category

    Societal Risks

    Political usage (Influencing Politics)

  49. 50.03.03 · Risk Sub-Category

    Societal Risks

    Political usage (Deterring democratic participation)

  50. 50.03.04 · Risk Sub-Category

    Societal Risks

    Political usage (Disrupting Social Order)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.