AI incident ·
Anthropic Report Details Claude Misuse for Influence Operations, Credential Stuffing, Recruitment Fraud, and Malware Development
In brief
An AI system built by Anthropic and deployed by Malicious Actors, Influence As A Service Operators and 1 other allegedly harmed Social Media Users, People Targeted By Malware and 5 others.
- Risk domain
- Malicious Actors & Misuse
- Occurred
- Coverage
- 5 reports
What happened
In April 2025, Anthropic published a report detailing several misuse cases involving its Claude LLM, all detected in March. These included an "influence-as-a-service" operation that orchestrated over 100 social media bots; an effort to scrape and test leaked credentials for security camera access; a recruitment fraud campaign targeting Eastern Europe; and a novice actor developing sophisticated malware. Anthropic banned the accounts involved but could not confirm downstream deployment.
Laws that address this harm
Policy angle: Classified under Malicious Actors & Misuse (Disinformation, surveillance, and influence at scale) in the MIT AI Risk Repository taxonomy; 5 recorded instruments address this use case.
- India DPDP Act
- Law on Artificial Intelligence (2025)
- Law No. 132/2025 on artificial intelligence
- EU AI Act
- Texas Responsible AI Governance Act (TRAIGA)
Matched from the record's risk domain and country to the instruments recorded here. A reviewer can correct the match in the repository (data/external/incident_overrides.yaml).
News reports (5)
Titles link to the original publisher; report text is not reproduced here.
Who was involved
- Alleged developer
- Anthropic
- Alleged harmed party
- Social Media Users, People Targeted By Malware, National Security And Intelligence Stakeholders, Job Seekers In Eastern Europe, Iot Security Camera Owners, Epistemic Integrity, Targets Of Fraudulent Professional Opportunities
Classification (MIT AI Risk Repository taxonomy)
- Risk domain
- Malicious Actors & Misuse
- Risk subdomain
- 4.1 Disinformation, surveillance, and influence at scale
- Causal entity
- Human
- Intent
- Intentional
- Timing
- Post-deployment
- Harm level
- —
- Sectors
- —
- Countries
- —
Risk entries describing this failure mode
Entries from the MIT AI Risk Repository coded to subdomain 4.1.
- Political manipulation
"Political manipulation - Use or misuse of personal data to target individuals’ interests, personalities and vulnerabilities with tailored political messages via micro-advertising or deepfakes/synthetic media."
- Coercion/manipulation
"Coercion/manipulation - Use of a technology system to covertly alter user beliefs and behaviour using nudging, dark patterns and/or other opaque techniques, resulting in potential erosion of privacy, addiction, anxiety/...
- Political and Economic
"Political and Economic - Manipulation of political beliefs, damage to political institutions and the effective delivery of government services."
- Electoral interference
"Electoral interference - Generation of false or misleading information that can interrupt or mislead voters and/or undermine trust in electoral processes."
- Political
"In the UK, a form of initial computational propaganda has already happened during the Brexit referendum1 . In future, there are concerns that oppressive governments could use AI to shape citizens’ opinions"
- Biased influence through citizen screening and tailored propaganda
"AI-powered chatbots tailor their communication approach to influence individual users' decisions. In the UK, a form of initial computational propaganda has already happened during the Brexit referendum. In future, there...
- Surveillance and Censorship
"Content moderation has emerged as one of the key use-cases of LLMs (Weng et al., 2023), indicating the potential of LLMs for surveillance and censorship as well (Edwards, 2023). Surveillance and censorship are one of th...
- Disinformation and manipulation of public opinion
"AI, particularly general- purpose AI, can be maliciously used for disinformation (351), which for the purpose of this report refers to false information that was generated or spread with the deliberate intent to mislead...
Incidents in the same risk subdomain
- Suspected AI-Generated Deepfake Video Reportedly Targeted Former Chhattisgarh Chief Minister Bhupesh Baghel on Instagram
- Purported AI-Generated War Footage Reportedly Circulated Widely Online During the Opening Phase of the War in Iran
- Purportedly AI-Manipulated Satellite Image Reportedly Claimed Iranian Strike Destroyed U.S. Radar in Qatar
- Network of Allegedly Fake Facebook Profiles with Purportedly AI-Generated Images Amplified Posts by Bulgaria's 'There Is Such a People' (ITN) Party
- Purportedly AI-Generated Image Reportedly Circulated Ahead of Thai Election Depicting PM Anutin Charnvirakul Dining with Benjamin Mauerberger
- Purportedly AI-Altered Images Reportedly Distort Evidence After Minneapolis Shooting of ICU Nurse Alex Pretti
Source record: incident #1054 on the AI Incident Database · all 5 reports