AI incident #377 ·

Weibo Model Had Difficulty Detecting Shifts in Censored Speech

What happened

Weibo's user moderation model is having difficulty keeping up with shifting user slang in defiance of Chinese state censors.

Only the incident metadata is stored here. The underlying news reports are on the AI Incident Database (CC BY-SA 4.0); use the links above to read them.

News reports (1)

Coverage catalogued by the AI Incident Database. Titles link to the original publisher; the text is not reproduced here.

  1. How Chinese citizens use puns to get past internet censors
    restofworld.org · Meaghan Tobin, Katherine Lee · AIID #2174

Who was involved

Alleged deployer
Weibo
Alleged developer
Weibo
Alleged harmed party
Weibo, Chinese Government

Classification (MIT AI Risk Repository taxonomy)

Causal entity
AI
Intent
Intentional
Timing
Post-deployment
Harm level
Sectors
Countries

Risk entries describing this failure mode

Entries from the MIT AI Risk Repository coded to subdomain 4.1.

  • Coercion/manipulation

    "Coercion/manipulation - Use of a technology system to covertly alter user beliefs and behaviour using nudging, dark patterns and/or other opaque techniques, resulting in potential erosion of privacy, addiction, anxiety/...

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Electoral interference

    "Electoral interference - Generation of false or misleading information that can interrupt or mislead voters and/or undermine trust in electoral processes."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Political manipulation

    "Political manipulation - Use or misuse of personal data to target individuals’ interests, personalities and vulnerabilities with tailored political messages via micro-advertising or deepfakes/synthetic media."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Political and Economic

    "Political and Economic - Manipulation of political beliefs, damage to political institutions and the effective delivery of government services."

    A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms (Abercrombie2024)

  • Biased influence through citizen screening and tailored propaganda

    "AI-powered chatbots tailor their communication approach to influence individual users' decisions. In the UK, a form of initial computational propaganda has already happened during the Brexit referendum. In future, there...

    The Rise of Artificial Intelligence - Future Outlooks and Emerging Risks (Allianz2018)

  • Political

    "In the UK, a form of initial computational propaganda has already happened during the Brexit referendum1 . In future, there are concerns that oppressive governments could use AI to shape citizens’ opinions"

    The Rise of Artificial Intelligence - Future Outlooks and Emerging Risks (Allianz2018)

  • Surveillance and Censorship

    "Content moderation has emerged as one of the key use-cases of LLMs (Weng et al., 2023), indicating the potential of LLMs for surveillance and censorship as well (Edwards, 2023). Surveillance and censorship are one of th...

    Foundational Challenges in Assuring Alignment and Safety of Large Language Models (Anwar2024)

  • Disinformation and manipulation of public opinion

    "AI, particularly general- purpose AI, can be maliciously used for disinformation (351), which for the purpose of this report refers to false information that was generated or spread with the deliberate intent to mislead...

    International Scientific Report on the Safety of Advanced AI (Bengio2024)

Incidents in the same risk subdomain

All incidents in this subdomain