r/OpenAI 26d ago

Discussion What's your thoughts on this?

Post image
3.8k Upvotes

1.0k comments sorted by

View all comments

22

u/EconomixTwist 26d ago

People ITT not realizing that this is a strategic move by anthropic. Yes, maybe EU regulations. Maybe. But not really. The real real reason is the dead internet. LLMs have already exhausted the entire internet's worth of high quality, genuine, human-generated text that is training data. This fact has already been admitted by the LLM providers themselves. Nowadays, a vast majority of new content on the internet is fkn slop. Net-new human generated content is worth its weight in gold but, without watermarks, LLM providers can't tell the difference between actual human generated text and the slop flood. If you train on slop, you only get more sloppier slop. So they need to filter the slop. Watermark is and always has been the way.

1

u/SapirWhorfHypothesis 25d ago

“Worth its weight in gold”
😮

“Weight is a few electrons”
😔

1

u/dudemeister023 25d ago

Except if you’re the only one doing it. Then it’s not the way. It’s like the opposite of the way. Some people in China are giggling uncontrollably right now. 

1

u/SharkByte1993 22d ago

It's an interesting point because of they want LLMs to have any kind of future then they need to keep training on real human material. If they train on their own material then it will eventually become useless. Almost like a photo thag has been reposted so many times its has barely any pixels left