this post was submitted on 17 Feb 2024
299 points (97.8% liked)

Technology

59605 readers
3394 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS
 

Reddit has a new AI training deal to sell user content::Reddit has reportedly made a deal with an unnamed AI company to allow access to its platform’s content for the purposes of AI model training.

you are viewing a single comment's thread
view the rest of the comments
[–] lvxferre@mander.xyz 9 points 9 months ago* (last edited 9 months ago) (2 children)

For anyone looking for a gibberish generator to replace their Reddit content with, here's one. This shit is like poison for those large models.

For automatic edition I'm not sure on what people can use nowadays; back then just before the APIcalypse I've used power delete suite, I'm not sure if it still works and I'm not creating a Reddit account just to test it out.

[–] greaprr@sh.itjust.works 0 points 9 months ago (1 children)

Not that I’m against telling Reddit to fuck off in no uncertain terms, but won’t providing this kind of poisoning to AI training just make it more resilient to exactly this kind of thing?

[–] lvxferre@mander.xyz 1 points 9 months ago* (last edited 9 months ago)

I don't think so. It's really hard to sort the poison out of the data, unless you actually have enough reading comprehension to know that it's gibberish - humans do, bots don't. And even if they discard 80% of the poison, the 20% there are already screwing with the model.

They could prevent you from editing your posts/comments, but that would cause an uproar.