this post was submitted on 28 Aug 2026
884 points (98.7% liked)

Technology

87624 readers
3575 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
 

For her safety, Doe has opted to receive alerts from the US Department of Justice Victim Notification System any time she may be a victim in a new criminal investigation. Although she has received countless alerts, she was shocked when the CCCP notified her that it had identified AI-generated CSAM on xAI that depicted her. This re-traumatized Doe, whose complaint alleged that messages were found on online forums “between offenders chatting about creating AI generated CSAM of Plaintiff and other similarly situated known, legacy, victims of CSAM.”

Now, Doe fears that xAI has not only made it easier to make more violative images of the most distressing time in her life, but also that xAI allegedly has stored the images that Grok generates and uses those outputs to further train Grok. Because of this, she believes that Grok has been trained on both the initial set of images that have haunted her for more than 20 years and the more recent AI-generated ones.

This is the first case to accuse xAI of training on CSAM, and the complaint does not go into great detail on that claim. Previously, Ars reported on a controversial dataset that was later scrubbed after researchers found CSAM in the training data, but there’s no indication xAI trained on that data. In a press release from lawyers representing Doe, it explained that Doe’s images were included in a CSAM Hash List maintained by NCMEC, and “that same material” allegedly “was part of the dataset xAI used to build Grok’s image and video generating capabilities.” The complaint similarly only alleged that “CSAM depicting Plaintiff with its longstanding well-known hash values has been used as a part of the dataset used by xAI.”

you are viewing a single comment's thread
view the rest of the comments
[–] mechoman444@lemmy.world 4 points 14 hours ago* (last edited 14 hours ago) (1 children)

What? That's not even remotely how that works.

A human being is also shaped by the information they've encountered throughout their life. You've encountered racism, stupidity, misinformation, and countless other forms of bad information. That doesn't mean you automatically become racist, stupid, or incapable of distinguishing good information from bad information.

AI works similarly in the sense that its training data contains an enormous amount of contradictory, inaccurate, biased, and outright terrible information. The mere presence of that information in the training data does not logically imply that the resulting model possesses those characteristics.

And AI isn't "dumb" or "smart" in the way you're describing. Those are human cognitive attributes. The relevant question is what a model can actually do, what patterns it has learned, and how reliably it performs a given task.

You don't like AI. Fine. But you clearly don't understand how it works, and this is an unfortunately constant problem on this platform. People make an assumption about AI, then present that assumption as though they've discovered some fundamental truth about the technology.

Misinformation is misinformation. It doesn't become correct because you think the conclusion is morally justified. You didn't figure out anything about AI here. You made an assumption and called it true.

All, I'm saying if you're going to critique something, at least know what you're critiquing or criticizing.

Also before anybody makes any more assumptions, what happen here is absolutely mind-blowingly horrific. Using csam to train anything is just terrible and exploitive. I'm very sorry for doe for having to continuously go through this.

Also also the people responsible for this should be charged with possession of CSAM and prosecuted in a court of law.

Disclaimer: Disagreeing with someone's argument, pointing out flaws in their reasoning, or interpreting their position differently does not automatically make it a straw man. For example, if I say, "I think we should reduce military spending," and you respond, "You think we should completely eliminate the military," you have created a straw man. You are arguing against a position I never actually took.

[–] Rothe@piefed.social 3 points 9 hours ago* (last edited 9 hours ago) (1 children)

The mere presence of that information in the training data does not logically imply that the resulting model possesses those characteristics.

Recurring studies do tend to show that AI does indeed posses such characteristics though.

[–] mechoman444@lemmy.world 3 points 6 hours ago

This source doesn't establish what you think it establishes.

First, it's a September 2021 infographic, updated in February 2024, not some contemporary study demonstrating that "AI in general" is racially prejudiced. More importantly, the subject here is health-care algorithms, and the article is specifically discussing algorithms that were deliberately designed to use race as a variable or that learned disparities from historical health-care data.

In fact, the source explicitly says that these systems can unintentionally increase existing racial biases through the explicit use of race in predicting outcomes and risk. That's not evidence that AI possesses racial prejudice. It's evidence that humans designed algorithms using race as a predictive variable, despite race being a poor proxy for genetic differences. The article even gives examples of medical algorithms where researchers subsequently removed race from the calculation.

That's an extremely important distinction you're completely glossing over.

If I build an algorithm that says "Black = higher risk" and the algorithm consequently produces a racial disparity, I've demonstrated that my algorithm contains a problematic racial assumption. I have not demonstrated that "AI is inherently racist." Likewise, if an algorithm uses health-care spending as a proxy for how sick someone is, and that proxy reflects existing racial disparities in access to health care, the resulting bias comes from the data and the proxy, not some intrinsic racial prejudice possessed by the AI.

And your source actually undermines the broader claim you're trying to make. It explicitly discusses AI being used to reduce racial disparities and cites research where algorithmic approaches improved outcomes or reduced unexplained disparities.

So yes, algorithmic bias in health care is a real and well-documented problem. Nobody is disputing that. What you're doing is taking evidence that specific algorithms can encode or reproduce human biases and extrapolating it into "AI itself is racially prejudiced."

That's not what your source says, and it isn't what the evidence demonstrates.