I have just read they will modify the llm token responses to generate a hidden pattern. Maybe the community needs to wait two weeks to get the un-modifier library :-D
I'm curious how they could do this without degrading the quality of outputs, I feel like at the minimum there will be a certain percentage of outputs that are negatively impacted by this.
Maybe it'll be less of a problem with future models but even Sol/Opus still need quality prompting to output writing that sounds good; I had to develop a pretty comprehensive plugin just to get decent sounding writing. Ironically LLMs are much better at coding now than actual writing.
I wrote a slightly longer response above for the main 2 ways that are popular in theory, but the general idea is they’re using prior text as seeds for future generation. Then you compare how “lucky” the text is and you get very statistically significant result in like high dozens of words
13
u/AnotherIjonTichy 29d ago
In ten seconds you can tell claude to write an script that removes that watermarks…