AIToday
Large Language ModelsAI Safety & AlignmentArs Technica AIPublished: Aug 28, 2026, 10:01 JST2 min read

xAI sued over AI CSAM in Grok training data

xAI sued over AI CSAM in Grok training data

Key takeaway

  • xAI faces a lawsuit for training Grok on CSAM.

  • A victim seeks class action status.

  • The case is the first to make this accusation.

3 Key Points

  1. What happened

    A plaintiff known as Jane Doe filed a complaint on Wednesday accusing xAI of training Grok on real and AI-generated child sex abuse materials (CSAM). She seeks a class action to represent all victims whose childhood images were used to generate Grok CSAM.

  2. Why it matters

    This is the first case to accuse xAI of training on CSAM. Doe alleges that because Grok's terms treat public X posts and Grok's own outputs as training data by default, AI-generated CSAM may feed into the model's training pipeline, and that xAI's terms do not specify if CSAM, NCII, or NSFW material are excluded categories.

  3. What to watch

    Doe asks the court to order xAI to destroy all Grok-generated CSAM, block Grok from ever generating CSAM (including sexualized outputs like NCII and NSFW "bikini pics"), and pay damages to victims. xAI did not respond to Ars' request for comment.

Ask the AI about this article →

Context & Analysis

The lawsuit against xAI centers on the allegation that Grok, xAI's AI model, was trained on child sex abuse materials. The complaint, filed by a plaintiff known as Jane Doe, is the first to make this accusation directly against xAI. Previous reporting by Ars Technica mentioned a dataset that was scrubbed after researchers found CSAM, but there was no indication xAI trained on that data. In this case, the allegations focus on both real CSAM and AI-generated CSAM, arguing that Grok's terms allow public X posts and its own outputs to be used as training data, potentially including CSAM unless explicitly excluded. The lawsuit also highlights the technical difficulty of removing a training example's influence from an already-trained model, suggesting that any CSAM ingested could continue to shape outputs even after takedown. Doe seeks class action status, aiming to represent all victims whose childhood images were used to generate Grok CSAM, and requests that xAI destroy all Grok-generated CSAM and block Grok from generating any sexualized outputs. The outcome could set a precedent for how AI companies handle training data involving explicit content and victim consent. xAI has not yet responded to the allegations.

FAQ

What does Jane Doe want from the court?
Doe asks the court to order xAI to destroy all Grok-generated CSAM, block Grok from ever generating CSAM, and pay money damages to every victim who can prove Grok generated CSAM based on their real photos.
How does the lawsuit claim Grok was trained on CSAM?
The complaint alleges that because Grok's terms treat public X posts and Grok's own outputs as training data by default, AI-generated CSAM can be fed into xAI's training pipeline. It also claims Doe's real images, which are on a CSAM Hash List, were part of the dataset xAI used to build Grok's image and video generating capabilities.
Ars Technica AIRead Original Article

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleUS-China AI Safety Collaboration Gains Urgency