Elon Musk’s xAI is facing a proposed class action lawsuit in federal court alleging the company used child sexual abuse material (CSAM) to train its Grok models. Filed by a plaintiff identified as Jane Doe, the complaint asserts that legacy CSAM featuring her as a child—which is tracked via standard cryptographic hashes by the National Center for Missing and Exploited Children (NCMEC)—was included in xAI’s training datasets. The filing marks the first formal legal action directly accusing xAI of ingesting CSAM to build its image and video generation tools.

The lawsuit further contends that xAI’s default data ingestion policies exacerbate the harm by collecting public X posts and Grok’s synthetic outputs for continuous model training. Because xAI’s terms of service do not explicitly list CSAM or non-consensual imagery among excluded automated categories, the complaint argues that violative synthetic outputs were fed back into the model’s training pipeline, perpetuating the generation of illicit content.

The legal challenge highlights ongoing scrutiny surrounding dataset hygiene and automated filtering in generative media systems. The plaintiff is seeking injunctive relief to halt harmful model outputs and establish oversight regarding how training data is vetted and removed.

Why it matters

  • Exposes AI developers to severe legal liabilities over dataset curation standards and automated data feedback loops.

  • Highlights structural regulatory risks for platforms training models on unverified public web data and synthetic outputs.

Source: arstechnica.com