this post was submitted on 14 Aug 2026
226 points (97.9% liked)

Technology

87184 readers
3655 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] gwheel@lemmy.zip 59 points 19 hours ago (2 children)

Assuming a checker tool is public it's enough for a teacher to verify classwork, but AI providers having the sole ability to identify generated content with no way to independently verify is not a real solution.

Plus this site advertises a tool to remove this watermarking, so it can't be that hard to scrub out if you're aware of it.

[–] Dojan@pawb.social 44 points 18 hours ago (1 children)

The goal is to ensure that they don’t inbreed their models, not fix the problems they’ve caused.

[–] brucethemoose@lemmy.world 21 points 18 hours ago* (last edited 18 hours ago)

This won't fix the inbreeding issue, anyway. The bias is extremely slight, but random, and orthogonal to Claude's own "slop patterns" and tendencies. And theres tons of other LLM content that will end up in their dataset outside their control.

Besides, as much as Claude accusess others of it, everyone's training on everyone else's output and they know it.

[–] Diurnambule@jlai.lu 5 points 18 hours ago (1 children)

I wonder what would happen if some start to watermark document they doesn't want in Claude training

[–] turtlesareneat@piefed.ca 3 points 14 hours ago

There's a key involved that we don't have, so people can't do this on their own. It's pretty fascinating. Training models would have to be told to check for watermarks and ignore them, but yeah that would be an effective way for the providers to avoid ingesting their own AI output.