Earlier quoted context omitted.
I'll be interested to see if people can actually pick out which text is watermarked and which isn't once they introduce it. It won't switch "bananas" to "airplanes". It'll switch "I really enjoy eating bananas" to "I love eating bananas" or similar
To that point, given a corpus of writing from person A and another from person B, I have no doubt it’s easy to train a classifier to determine who wrote what. In fact I believe law enforcement agencies already have these classifiers. The only difference here is that Anthropic is actively trying to make the watermark undetectable.
Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
281–290 of 776 posts
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#282My biggest concern is that checking any text for watermarks requires sending the entire text to Anthropic. And even that is not sufficient, as the text might have been generated with ChatGPT, Gemini, Grok, Mistral, ... So every check requires sending the text to as many AI providers as offer a watermarking detection API, almost all of which have a very dubious track history with obtaining training data through illici…
> blindly trusting they won't train on any of that being allowed to train on any data that you can legally obtain ought to be a right for anyone. After all, i am allowed to learn off anything i can legally read (and perhaps even illegally read). The only thing not allowed (rightly so) is to produce a copy with enough similarities that it can be replacing the original.
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#283Earlier quoted context omitted.
I think the author is just mad people will be able to detect and filter out their AI slop writing in the future.
This is not it, no. He is not using AI and it is not I think remotely in his nature to surrender that control. He is engaging with this on principle. Again I am not sure I agree with him, but then it’s a hypothetical because I am not going to get an LLM to write for me either.
That's ... even worse? So we're all here in the comments trying to figure out what the author means, and what their overall point is, while clearly they don't even use the damn thing? Oof... What a waste of time for everyone involved.
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#284Earlier quoted context omitted.
> blindly trusting they won't train on any of that being allowed to train on any data that you can legally obtain ought to be a right for anyone. After all, i am allowed to learn off anything i can legally read (and perhaps even illegally read). The only thing not allowed (rightly so) is to produce a copy with enough similarities that it can be replacing the original.
Why would that be a legal right? Why should we hand over even MORE power to the owner class? In a fantasy world this could be possible yes.
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#285Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#286Gruber shows here that he really doesn’t understand the basics of how LLM text generation works. It’s weird he picked this battle about the quality of writing in LLMs. Was he planning to use LLMs to write his articles? Well, not that weird actually. He just has a hard-on against anything that comes from the EU since Apple got in trouble. If the EU said tomorrow that they want peace in the world he’d be in Fox News th…
I dunno, I guess that's what you should expect from Gruber but these EU-bashing articles lowered the enjoyment I got from his blog underneath the bar for me.
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#287Earlier quoted context omitted.
> blindly trusting they won't train on any of that being allowed to train on any data that you can legally obtain ought to be a right for anyone. After all, i am allowed to learn off anything i can legally read (and perhaps even illegally read). The only thing not allowed (rightly so) is to produce a copy with enough similarities that it can be replacing the original.
Why would that be a legal right? Why should we hand over even MORE power to the owner class? In a fantasy world this could be possible yes.
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#288Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#289Earlier quoted context omitted.
> blindly trusting they won't train on any of that being allowed to train on any data that you can legally obtain ought to be a right for anyone. After all, i am allowed to learn off anything i can legally read (and perhaps even illegally read). The only thing not allowed (rightly so) is to produce a copy with enough similarities that it can be replacing the original.
Why would that be a legal right? Why should we hand over even MORE power to the owner class? In a fantasy world this could be possible yes.
Re: Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
#290Earlier quoted context omitted.
To that point, given a corpus of writing from person A and another from person B, I have no doubt it’s easy to train a classifier to determine who wrote what. In fact I believe law enforcement agencies already have these classifiers. The only difference here is that Anthropic is actively trying to make the watermark undetectable.
They're not trying to make the watermark undetectable, that would defeat the point of a watermark. They're making it detectable, but not make the text obviously watermarked