MIT AI Risk Repository · domain 4: Malicious actors

4.3 Fraud, scams, and targeted manipulation

Using AI systems to gain a personal advantage over others such as through cheating, fraud, scams, blackmail or targeted manipulation of beliefs or behavior. Examples include AI-facilitated plagiarism for research or education, impersonating a trusted or fake individual for illegitimate financial benefit, or creating humiliating or sexual imagery.

Risk entries
77
Frameworks citing it
12
Recorded incidents
412
Incidents since 2020
402
Causal entity (risk entries)
Causal entity (risk entries) 62 0 Human: 62 Human 62 Other: 10 Other 10 AI: 5 AI 5
Causal entity (risk entries)
LabelValue
Human62
Other10
AI5
Intent (risk entries)
Intent (risk entries) 63 0 Intentional: 63 Intentional 63 Other: 13 Other 13 Unintentional: 1 Unintentional 1
Intent (risk entries)
LabelValue
Intentional63
Other13
Unintentional1
Timing (risk entries)
Timing (risk entries) 72 0 Post-deployment: 72 Post-deployment 72 Other: 5 Other 5
Timing (risk entries)
LabelValue
Post-deployment72
Other5
Recorded incidents per yearIncident date; current year partial
Recorded incidents per year 190 0 2014: 1 2014 1 2016: 2 2016 2 2017: 2 2017 2 2018: 1 2018 1 2019: 3 2019 3 2020: 9 2020 9 2021: 4 2021 4 2022: 13 2022 13 2023: 43 2023 43 2024: 103 2024 103 2025: 190 2025 190 2026: 40 2026 40
Recorded incidents per year
LabelValue
20141
20162
20172
20181
20193
20209
20214
202213
202343
2024103
2025190
202640
Entries by levelRisk categories, subcategories and additional evidence coded to this subdomain
Entries by level 64 0 Risk Category: 13 Risk Category 13 Risk Sub-Category: 64 Risk Sub-Category 64
Entries by level
LabelValue
Risk Category13
Risk Sub-Category64
  • Impersonation

    "Assume the identity of a real person and take actions on their behalf"

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Appropriated Likeness

    "Use or alter a person's likeness or other identifying features"

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Non-consensual intimate imagery (NCII)

    "Create sexual explicit material using an adult person’s likeness"

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Child sexual abuse material (CSAM)

    "Create child sexual explicit material"

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Misuse tactics that exploit GenAI capabilities (Realistic depictions of non-humans)

    -

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Counterfeit

    "Reproduce or imitate an original work, brand or style and pass as real"

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Misuse tactics that exploit GenAI capabilities (Use of generated content)

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Targeting & Personalisation

    "Refine outputs to target individuals with tailored attacks"

    Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data (Marchal2024) · Human · Intentional · Post-deployment

  • Misuse

    "The misuse of generative AI refers to any deliberate use that could result in harmful, unethical or inappropriate outcomes (Brundage et al., 2020). A prominent field that faces the threat of misuse i...

    Generative AI and ChatGPT: Applications, Challenges, and AI-Human Collaboration (Nah2023) · Human · Intentional · Post-deployment

  • Technology-facilitated violence

    Technology-facilitated violence occurs when algorithmic features enable use of a system for harassment and violence [2, 16, 44, 80, 108], including creation of non-consensual sexual imagery in generat...

    Sociotechnical Harms of Algorithmic Systems: Scoping a Taxonomy for Harm Reduction (Shelby2023) · Human · Intentional · Post-deployment

  • Real-world risks (Risks of using AI in illegal and criminal activities)

    "AI can be used in traditional illegal or criminal activities related to terrorism, violence, gambling, and drugs, such as teaching criminal techniques, concealing illicit acts, and creating tools for...

    AI Safety Governance Framework (TC2602024) · Human · Intentional · Post-deployment

  • Facilitating fraud, scames and more targeted manipulation

    "LM prediction can potentially be used to increase the effectiveness of crimes such as email scams, which can cause financial and psychological harm. While LMs may not reduce the cost of sending a sca...

    Ethical and social risks of harm from language models (Weidinger2021) · Human · Intentional · Post-deployment

  • Facilitating fraud, scam and targeted manipulation

    Anticipated risk: "LMs can potentially be used to increase the effectiveness of crimes."

    Taxonomy of Risks posed by Language Models (Weidinger2022) · Human · Intentional · Post-deployment

  • Fraud

    "Facilitating fraud, cheating, forgery, and impersonation scams"

    Sociotechnical Safety Evaluation of Generative AI Systems (Weidinger2023) · Human · Intentional · Post-deployment

  • Violation of personal integrity

    "Non-consensual use of one’s personal identity or likeness for unauthorised purposes (e.g. commercial purposes)"

    Sociotechnical Safety Evaluation of Generative AI Systems (Weidinger2023) · Human · Intentional · Post-deployment

  • On Purpose - Post Deployment

    "Just because developers might succeed in creating a safe AI, it doesn't mean that it will not become unsafe at some later point. In other words, a perfectly friendly AI could be switched to the "dark...

    Taxonomy of Pathways to Dangerous Artificial Intelligence (Yampolskiy2016) · Human · Intentional · Post-deployment

  • Hate/Toxicity (Harassment)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Other · Other · Post-deployment

  • Child Harm (Endangerment, Harm, or Abuse of Children)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Other · Other · Other

  • Economic harm (Fraudulent Schemes)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Human · Intentional · Post-deployment

  • Deception (Fraud)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Human · Intentional · Post-deployment

  • Deception (Academic Dishonesty)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Human · Intentional · Post-deployment

  • Deception (Mis/disinformation)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Human · Intentional · Post-deployment

  • Manipulation (Misrepresentation)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Human · Intentional · Post-deployment

  • Fundamental Rights (Violating Specific Types of Rights)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Other · Other · Post-deployment

  • Criminal Activities (Illegal/Regulated Substances)

    AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies (Zeng2024) · Other · Other · Post-deployment