Earlier quoted context omitted.
It was quick :) … https://claudewatermarkremover.app/
> Honest note: Anthropic has not shipped a public Claude watermark detector yet. This tool uses rewrite-based neutralization — a meaning-preserving paraphrase with a non-Claude model — which is the attack path watermark research points to. Not affiliated with Anthropic. Well, they should have run their own AI slop website through their tool...
How Claude marks AI-generated content
251–260 of 446 posts
Re: How Claude marks AI-generated content
#252Earlier quoted context omitted.
Thankfully, there are a variety of Chinese models that never will. I think we all know that in a few years, they will also be the only relevant offerings on the market, due to not being bogged down with over-zealous ""safety"" footguns.
Your theory is that the Chinese government is thoroughly uninterested in safety or prosocial controls?
Re: How Claude marks AI-generated content
#253Earlier quoted context omitted.
My guess is it will be similar to how Genius watermarked lyrics, using things like variants of punctuation https://www.pcmag.com/news/genius-we-caught-google-red-hande...
In program code? Unlikely, surely¡
Re: How Claude marks AI-generated content
#254The SV obsession with neo Kabbalistic nonsense will get a whole new burst of energy from this.
Re: How Claude marks AI-generated content
#255Re: How Claude marks AI-generated content
#256We need to just stop pretending we can reliably tell if plain text is written by an LLM. It’s just not a reasonable ask.
Re: How Claude marks AI-generated content
#257>Claude may not be the original author. People often use Claude to proofread, translate, summarize, or convert files. The output can carry a Claude mark even if the underlying ideas, text, or data originated from another source;
Such models already struggle not making any unnecessary or unwanted changes to a corpus, this makes them unable to by design.
Re: How Claude marks AI-generated content
#258Earlier quoted context omitted.
They aren't using greedy decoding, there's enough randomness in sampling to swap some with independent signal.
Purely greedy or not, there is some measure of "goal outcome" that was previously being solved for with the token selection function, and the goal was "complete this text with the best (surely, otherwise what are we doing?) next part, and sometimes the best next part is a little bit random just to keep things interesting" Now the goal is either "identify the meaningless interesting bits and swap them out with 0% loss…
Re: How Claude marks AI-generated content
#259Earlier quoted context omitted.
Your theory is that the Chinese government is thoroughly uninterested in safety or prosocial controls?
Amongst Chinese labs and netizens, there's MUCH less belief/mindshare on "AGI = existential risk to humanity", "paperclip maximiser", and similar lines of thinking. AI is seen more as just a technology, and less like a scary boogyman. Whether that's right or wrong, I'll leave to you, but there's huge differences in perspectives, and if you only get your news from Western sources and communities (and companies), you'r…
For a taste of where I think things are headed, try asking Chinese models about Tiananmen [1]. And then take a look at the Chinese government's approach to pretty much anything that they think reduces security or social harmony. I find it hard to believe their models will be the one exception to that over the long term.
[1] https://en.wikipedia.org/wiki/1989_Tiananmen_Square_protests...
Re: How Claude marks AI-generated content
#260Earlier quoted context omitted.
Your "code that Claude makes..."? Oh, how I laughed. That was never your code, my friend.
Please don't let the arbitrary selection of phrase distract you from the substance of my argument: a product that I pay for is at best no better due to this change, and highly probably worse. Why am I paying for a tool that is beholden to clandestinely satisfy some far away master?
Why were you doing that before watermarking?
Same answer.