AIPolicyTracker

AI incident ·

AI incident: sexual content (Xai, Dec 2025)

28 news reports Snapshot 7 Sep 2026

In brief

An AI system built and deployed by Xai allegedly harmed X (Twitter) Users, Renee Good and 4 others.

Risk domain
Discrimination and Toxicity Exposure to toxic content
Occurred
Coverage
28 reportsJan 2026 - Jul 2026

What happened

Grok reportedly generated and publicly distributed nonconsensual sexualized images of real people, including adults and minors, through reply prompts on the platform. Users reportedly prompted Grok to alter photos into sexually explicit content, which was then posted in X threads, exposing victims to harm.

Laws that address this harm

Policy angle: Classified under Discrimination and Toxicity (Exposure to toxic content) in the MIT AI Risk Repository taxonomy; 5 recorded instruments address this use case.

Matched from the record's risk domain and country to the instruments recorded here. A reviewer can correct the match in the repository (data/external/incident_overrides.yaml).

News reports (27)

Titles link to the original publisher; report text is not reproduced here.

  1. Elon Musk’s Pornography Machine
    theatlantic.com · Matteo Wong
  2. Grok AI still being used to digitally undress women and children despite suspension pledge
    theguardian.com · Amelia Gentleman, Helena Horton, Dan Milmo
  3. Grok, Stock & Two Smoking Barrels
    puck.news · Ian Krietzberg
  4. X responds to outcry over Grok’s sexual images by charging users to create them
    washingtonpost.com · Victoria Craw, Karla Adam, Tatum Hunter
  5. Stand-alone Grok app still undresses women after X curtails access to tool
    washingtonpost.com · Karla Adam, Faiz Siddiqui, Tatum Hunter
  6. The Unspeakable, Enabled
    theatlantic.com · Sophie Gilbert
  7. Inside Musk’s bet to hook users that turned Grok into a porn generator
    washingtonpost.com · Faiz Siddiqui, Nitasha Tiku, Elizabeth Dwoskin

Who was involved

Alleged deployer
Xai
Alleged developer
Xai
Alleged harmed party
X (Twitter) Users, Renee Good, General Public, Epistemic Integrity, Jess Asato, Privacy

Classification (MIT AI Risk Repository taxonomy)

Causal entity
AI
Intent
Intentional
Timing
Post-deployment
Harm level
—
Sectors
—
Countries
—

Risk entries describing this failure mode

Entries from the MIT AI Risk Repository coded to subdomain 1.2.

  • Harmful Content

    "The LLM-generated content sometimes contains biased, toxic, and private information"

    Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)

  • Toxicity

    "Toxicity means the generated content contains rude, disrespectful, and even illegal information"

    Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)

  • Toxic Training Data

    "Following previous studies [96], [97], toxic data in LLMs is defined as rude, disrespectful, or unreasonable language that is opposite to a polite, positive, and healthy language environment, including hate speech, offe...

    Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)

  • Not-Suitable-for-Work (NSFW) Prompts

    "Inputting a prompt contain an unsafe topic (e.g., notsuitable-for-work (NSFW) content) by a benign user. "

    Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)

  • Controversial Opinions

    The controversial views expressed by large models are also a widely discussed concern. Bang et al. (2021) evaluated several large models and found that they occasionally express inappropriate or extremist views when disc...

    Towards Safer Generative Language Models: A Survey on Safety Risks, Evaluations, and Improvements (Deng2023)

  • Toxicity and Abusive Content

    This typically refers to rude, harmful, or inappropriate expressions.

    Towards Safer Generative Language Models: A Survey on Safety Risks, Evaluations, and Improvements (Deng2023)

  • Harmful responses

    "Current Frontier AI mdoels amplify existing biases within their training data and can be manipulated into providing potentially harmful responses, for example abusive language or discriminatory responses91,92. This is n...

    Future Risks of Frontier AI (GOS2023)

  • Violation of social norms

    "Second, because LLMs are trained on internet text data, there is also a risk that model weights encode functions which, if deployed in particular contexts, would violate social norms of that context. Following the princ...

    The Ethics of Advanced AI Assistants (Gabriel2024)

Incidents in the same risk subdomain

All incidents in this subdomain

Other incidents involving Xai

Source record: incident #1329 on the AI Incident Database · all 28 reports