this post was submitted on 29 Aug 2026
107 points (98.2% liked)

Technology

7204 readers
338 users here now

News community around technology, social media platforms, information technology and governmental policy surrounding it.

What doesn't fit here?

The core of the story has to be technology focused.


Post guidelines

Title formatPost title should mirror the news source title. If you don't like the title of article, look for an alternative source instead of editorializing it.
URL formatPost URL should be the original link to the article (even if paywalled) and archived copies left in the body. It allows avoiding duplicate posts when cross-posting.
[Opinion] prefixOpinion (op-ed) articles must use [Opinion] prefix before the title. Opinion articles refer to articles that their publisher doesn't explictly endorse.
Country prefixCountry prefix can be added to the title with a separator (|, :, etc.) if the news is from a local publisher who doesn't clearly mention the country.


Rules

1. English onlyTitle and associated content has to be in English.
2. Use original linkPost URL should be the original link to the article (even if paywalled) and archived copies left in the body. It allows avoiding duplicate posts when cross-posting.
3. Respectful communicationAll communication has to be respectful of differing opinions, viewpoints, and experiences.
4. InclusivityEveryone is welcome here regardless of age, body size, visible or invisible disability, ethnicity, sex characteristics, gender identity and expression, education, socio-economic status, nationality, personal appearance, race, caste, color, religion, or sexual identity and orientation.
5. Ad hominem attacksAny kind of personal attacks are expressly forbidden. If you can't argue your position without attacking a person's character, you already lost the argument.
6. Off-topic tangentsStay on topic. Keep it relevant.
7. Instance rules may applyIf something is not covered by community rules, but are against lemmy.zip instance rules, they will be enforced.


Companion communities

!globalnews@lemmy.zip
!interestingshare@lemmy.zip


Icon attribution | Banner attribution


If someone is interested in moderating this community, message @brikox@lemmy.zip.

founded 2 years ago
MODERATORS
top 9 comments
sorted by: hot top controversial new old
[–] NaibofTabr@infosec.pub 7 points 9 hours ago* (last edited 9 hours ago) (2 children)

I just want to point out that there is a valid use case for such a model. The massive amount of images uploaded to the internet every hour has created an exponentially growing need for content moderation, and unfortunately that means some people have to spend time reviewing material flagged as CSAM. This leads to a lot of mental trauma: https://www.researchgate.net/publication/373947227_The_psychological_impacts_of_content_moderation_on_content_moderators_A_qualitative_study (not surprisingly).

This is a fucking awful job, and there's way too much for human moderators to handle. Most people who do this job burn out within a year or two. Supporting their work with automated recognition takes some of the strain off of them. In an ideal world we wouldn't need humans to do this work at all, but that's unlikely. This is a very rare case where I 100% support replacing human labor with machine learning as much as possible. I can't really imagine a worse job, personally.

There is a real benefit to training some models to automatically identify CSAM, but such a model should be specialized and only used internally for its intended purpose. It should never be publicly accessible. I suspect that what happened with Grok is an example of gross negligence/incompetence, where a model intended for CSAM identification was included in the overall Grok system.

If that's not the explanation, then this is way more disturbing because anything else would have been done with intent.

[–] apotheotic@beehaw.org 3 points 4 hours ago* (last edited 4 hours ago)

There's a big difference between training Grok (their LLM and generative-ai offering) on it, and training a moderation tool

[–] pulsewidth@lemmy.world 14 points 8 hours ago* (last edited 8 hours ago)

Occam's Razor, I suspect X AI simply downloaded/scraped all porn they possibly could access on the internet to build their training data, and this included a large amount of child porn.

Likely Elon-style factors:

  • time pressure. performed in a huge rush (he joined the LLM party late with massive FOMO),
  • zero internal pushback. Majority highly pressured immigrant/temporary staff (H1-B visas used a lot by Twitter in wake of Musk buyout)
  • zero importance placed on ethics / content review by leadership, likely even an ethics-negative internal culture. "Move fast, break things, fuck the government regs".
  • loathe to pay licensing fees for materials or 'waste time' negotiating content-access contracts from legitimate porn businesses ("just scrape all the torrents and free sites you can find, the internet is full of free porn").
[–] cowfodder@lemmy.zip 17 points 12 hours ago

Mostly from Musk's personal collection.

[–] Tollana1234567@lemmy.today 5 points 10 hours ago

AS musk is a pedophile himself. he made it as king of the pedos in online spaces.

[–] CubitOom@infosec.pub 5 points 12 hours ago (3 children)
[–] sidebro@lemmy.zip 4 points 6 hours ago

Grok is a stupid enough name but I hate SpaceXAI even more

[–] floofloof@lemmy.ca 12 points 11 hours ago

So are half of his children.

[–] LadyMeow@lemmy.blahaj.zone 4 points 12 hours ago

Somehow even worse.