MIT AI Risk Repository · domain 4: Malicious actors

4.3 Fraud, scams, and targeted manipulation

Using AI systems to gain a personal advantage over others such as through cheating, fraud, scams, blackmail or targeted manipulation of beliefs or behavior. Examples include AI-facilitated plagiarism for research or education, impersonating a trusted or fake individual for illegitimate financial benefit, or creating humiliating or sexual imagery.

Risk entries
77
Frameworks citing it
12
Recorded incidents
412
Incidents since 2020
402
Causal entity (risk entries)
Causal entity (risk entries) 62 0 Human: 62 Human 62 Other: 10 Other 10 AI: 5 AI 5
Causal entity (risk entries)
LabelValue
Human62
Other10
AI5
Intent (risk entries)
Intent (risk entries) 63 0 Intentional: 63 Intentional 63 Other: 13 Other 13 Unintentional: 1 Unintentional 1
Intent (risk entries)
LabelValue
Intentional63
Other13
Unintentional1
Timing (risk entries)
Timing (risk entries) 72 0 Post-deployment: 72 Post-deployment 72 Other: 5 Other 5
Timing (risk entries)
LabelValue
Post-deployment72
Other5
Recorded incidents per yearIncident date; current year partial
Recorded incidents per year 190 0 2014: 1 2014 1 2016: 2 2016 2 2017: 2 2017 2 2018: 1 2018 1 2019: 3 2019 3 2020: 9 2020 9 2021: 4 2021 4 2022: 13 2022 13 2023: 43 2023 43 2024: 103 2024 103 2025: 190 2025 190 2026: 40 2026 40
Recorded incidents per year
LabelValue
20141
20162
20172
20181
20193
20209
20214
202213
202343
2024103
2025190
202640
Entries by levelRisk categories, subcategories and additional evidence coded to this subdomain
Entries by level 64 0 Risk Category: 13 Risk Category 13 Risk Sub-Category: 64 Risk Sub-Category 64
Entries by level
LabelValue
Risk Category13
Risk Sub-Category64
  • Impersonation/identity theft

    "Impersonation/identity theft - Theft of an individual, group or organisation’s identity by a third-party in order to defraud, mock or otherwise harm them."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024) · Human · Intentional · Post-deployment

  • IP/copyright loss

    "IP/copyright loss - Misuse or abuse of an individual or organisation’s intellectual property, including copyright, trademarks, and patents."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024) · Human · Intentional · Other

  • Dehumanisation/objectification

    "Dehumanisation/objectification - Use or misuse of a technology system to depict and/or treat people as not human, less than human, or as objects."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024) · Human · Other · Post-deployment

  • Defamation/libel/slander

    "Defamation/libel/slander - Use of a technology system to create, facilitate or amplify false perception(s) about an individual, group, or organisation."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024) · Human · Intentional · Post-deployment

  • Financial and business

    "Financial and Business - Use or misuse of a technology system in a manner that damages the financial interests of an individual or group, or which causes strategic, operational, legal or financial ha...

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024) · Human · Other · Post-deployment

  • Cheating/plagiarism

    "Cheating/plagiarism - Use of another person’s or group’s words or ideas without consent and/or acknowledgement."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024) · Other · Other · Other

  • Misinformation and Manipulation

    "Recent studies have demonstrated that LLMs can be exploited to craft deceptive narratives with levels of persuasiveness similar to human-generated content (Pan et al., 2023b; Spitale et al., 2023), t...

    Foundational Challenges in Assuring Alignment and Safety of Large Language Models (Anwar2024) · Human · Intentional · Post-deployment

  • Cybersecurity

    "LLMs may exacerbate cybersecurity risks in various ways (Newman, 2024). Firstly, LLMs may significantly amplify the effectiveness of deceptive operations aimed at tricking people into disclosing sens...

    Foundational Challenges in Assuring Alignment and Safety of Large Language Models (Anwar2024) · Human · Intentional · Post-deployment

  • Domain-Specific Misuses

    "Improvements in LLMs may exert greater pressure to apply LLMs to various domains, such as health and education (Eloundou et al., 2023). Crude efforts to use LLMs in such domains, however, may incur h...

    Foundational Challenges in Assuring Alignment and Safety of Large Language Models (Anwar2024) · Human · Intentional · Post-deployment

  • Harm to individuals through fake content

    "General- purpose AI systems can be used to increase the scale and sophistication of scams and fraud, for example through general- purpose AI- enhanced ‘phishing’ attacks. General- purpose AI can be u...

    International Scientific Report on the Safety of Advanced AI (Bengio2024) · Human · Intentional · Post-deployment

  • Harm to individuals through fake content

    "Malicious actors can use general- purpose AI to generate fake content that harms individuals in a targeted way. For example, they can use such fake content for scams, extortion, psychological manipul...

    International AI Safety Report 2025 (Bengio2025) · Human · Intentional · Post-deployment

  • Unhelpful Uses

    "Improper uses of LLM systems can cause adverse social impacts."

    Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024) · Human · Intentional · Post-deployment

  • Academic Misconduct

    "Improper use of LLM systems (i.e., abuse of LLM systems) will cause adverse social impacts, such as academic misconduct."

    Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024) · Human · Intentional · Post-deployment

  • Scams

    "Bad actors can also use generative AI tools to produce adaptable content designed to support a campaign, political agenda, or hateful position and spread that information quickly and inexpensively ac...

    Generating Harms - Generative AI's impact and paths forwards (EPIC2023) · Human · Intentional · Post-deployment

  • Harassment, Impersonation, and Extortion

    "Deepfakes and other AI-generated content can be used to facilitate or exacerbate many of the harms listed throughout this report, but this section focuses on one subset: intentional, targeted abuse o...

    Generating Harms - Generative AI's impact and paths forwards (EPIC2023) · Human · Intentional · Post-deployment

  • Malicious intent

    "A frequent malicious use case of generative AI to harm, humiliate, or sexualize another person involves generating deepfakes of nonconsensual sexual imagery or videos."

    Generating Harms - Generative AI's impact and paths forwards (EPIC2023) · Human · Intentional · Post-deployment

  • Privacy and consent

    "Even when a victim of targeted, AIgenerated harms successfully identifies a deepfake creator with malicious intent, they may still struggle to redress many harms because the generated image or video...

    Generating Harms - Generative AI's impact and paths forwards (EPIC2023) · Human · Intentional · Post-deployment

  • Believability

    Deepfakes can impose real social injuries on their subjects when they are circulated to viewers who think they are real. Even when a deepfake is debunked, it can have a persistent negative impact on h...

    Generating Harms - Generative AI's impact and paths forwards (EPIC2023) · Human · Intentional · Post-deployment

  • Data Security Risk

    "Just as every other type of individual and organization has explored possible use cases for generative AI products, so too have malicious actors. This could take the form of facilitating or scaling u...

    Generating Harms - Generative AI's impact and paths forwards (EPIC2023) · Human · Intentional · Other

  • Deception - Synthetic identities

    "GenAI can produce images of people that look very real, as if they could be seen on platforms like Facebook, Twitter, or Tinder. Although these individuals do not exist in reality, these synthetic id...

    GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023) · Human · Intentional · Post-deployment

  • Propaganda - Digital impersonations

    "AI-generated impersonation for identity theft might be found at the intersection of “Harm to the Person” and “Deception.”"

    GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023) · Other · Intentional · Post-deployment

  • Dishonesty - Targeted harassment

    "LLMs can be deployed to target individuals online, sending them personalized and harmful messages at scale"

    GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023) · Human · Intentional · Post-deployment

  • Financial and Economic Damage

    "Then, we have the potential for financial loss, fraud, market manipulation, and other economic harms, which fall under “Financial and Economic Damage.”

    GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023) · Other · Intentional · Post-deployment

  • Deception - Bespoke ransom

    -

    GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023) · Other · Intentional · Post-deployment

  • Dishonesty - Market manipulation

    -

    GenAI against humanity: nefarious applications of generative artificial intelligence and large language models (Ferrara2023) · Other · Intentional · Post-deployment