AIPolicyTracker

AI incident ·

AI-Powered Chinese Surveillance Campaign 'Peer Review' Used for Real-Time Monitoring of Anti-State Speech on Western Social Media

6 news reports Snapshot 7 Sep 2026

In brief

An AI system built by Various Open Source Ai Developers, Openai and 3 others and deployed by Chinese State Linked Actors and Chinese Communist Party allegedly harmed Western Social Media Communities, Social Media Users In Latin America and 5 others.

Risk domain
Malicious Actors & Misuse Disinformation, surveillance, and influence at scale
Occurred
Coverage
6 reportsFeb 2025

What happened

OpenAI reportedly uncovered evidence of a Chinese state-linked AI-powered surveillance campaign, dubbed "Peer Review," designed to monitor and report anti-state speech on Western social media in real time. The system, believed to be built on Meta’s open-source Llama model, was detected when a developer allegedly used OpenAI’s technology to debug its code. OpenAI also reportedly identified disinformation efforts targeting Chinese dissidents and spreading propaganda in Latin America.

Laws that address this harm

Policy angle: Classified under Malicious Actors & Misuse (Disinformation, surveillance, and influence at scale) in the MIT AI Risk Repository taxonomy; 5 recorded instruments address this use case.

Matched from the record's risk domain and country to the instruments recorded here. A reviewer can correct the match in the repository (data/external/incident_overrides.yaml).

News reports (6)

Titles link to the original publisher; report text is not reproduced here.

Who was involved

Alleged harmed party
Western Social Media Communities, Social Media Users In Latin America, Social Media Users, Opposition Voices Against The Chinese Communist Party, Chinese Dissidents, Cai Xia, Privacy

Classification (MIT AI Risk Repository taxonomy)

Causal entity
Human
Intent
Intentional
Timing
Post-deployment
Harm level
—
Sectors
—
Countries
—

Risk entries describing this failure mode

Entries from the MIT AI Risk Repository coded to subdomain 4.1.

  • Political manipulation

    "Political manipulation - Use or misuse of personal data to target individuals’ interests, personalities and vulnerabilities with tailored political messages via micro-advertising or deepfakes/synthetic media."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Coercion/manipulation

    "Coercion/manipulation - Use of a technology system to covertly alter user beliefs and behaviour using nudging, dark patterns and/or other opaque techniques, resulting in potential erosion of privacy, addiction, anxiety/...

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Political and Economic

    "Political and Economic - Manipulation of political beliefs, damage to political institutions and the effective delivery of government services."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Electoral interference

    "Electoral interference - Generation of false or misleading information that can interrupt or mislead voters and/or undermine trust in electoral processes."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Political

    "In the UK, a form of initial computational propaganda has already happened during the Brexit referendum1 . In future, there are concerns that oppressive governments could use AI to shape citizens’ opinions"

    The Rise of Artificial Intelligence - Future Outlooks and Emerging Risks (Allianz2018)

  • Biased influence through citizen screening and tailored propaganda

    "AI-powered chatbots tailor their communication approach to influence individual users' decisions. In the UK, a form of initial computational propaganda has already happened during the Brexit referendum. In future, there...

    The Rise of Artificial Intelligence - Future Outlooks and Emerging Risks (Allianz2018)

  • Surveillance and Censorship

    "Content moderation has emerged as one of the key use-cases of LLMs (Weng et al., 2023), indicating the potential of LLMs for surveillance and censorship as well (Edwards, 2023). Surveillance and censorship are one of th...

    Foundational Challenges in Assuring Alignment and Safety of Large Language Models (Anwar2024)

  • Disinformation and manipulation of public opinion

    "AI, particularly general- purpose AI, can be maliciously used for disinformation (351), which for the purpose of this report refers to false information that was generated or spread with the deliberate intent to mislead...

    International Scientific Report on the Safety of Advanced AI (Bengio2024)

Incidents in the same risk subdomain

All incidents in this subdomain

Source record: incident #939 on the AI Incident Database · all 6 reports