Live data from Hacker News

How Claude marks AI-generated content

support.claude.com

431–440 of 446 posts

Re: How Claude marks AI-generated content

#431

Earlier quoted context omitted.

Are you worried about being accused of using LLMs to generate your work? As long as you don't plagiarize you have nothing to worry about.

You can't make a blanket statement like this without knowing how the watermark is implemented.

Why not? I'm assuming that the watermark detector won't have plausible false positives on human-written text otherwise it won't have much merit to begin with.

If a detector flags something in your text and you've properly attributed that text to another author then what is there to worry about?

Re: How Claude marks AI-generated content

#433

So, does that watermark include attribution to the author of the original work used to train the model to begin with?

And. if you’re paying a subscription price or a per-token bill, doesn’t that effectively make any content or code generated under that subscription a “work for hire”?

Re: How Claude marks AI-generated content

#434
post #369

Earlier quoted context omitted.

It’s worse than that, false positives are possible but someone generating text should be able to get ai to change some words and formatting to break the watermarking, then ai detectors can tell them how well they did. I don’t know what the answer but I absolutely know it isn’t this.

I think we need a chain of custody system for content, but that would require browsers, software, websites, operating systems, phones, camera manufacturers, etc to all get on board. But each intermediary or source (optionally) cryptographicaly signs a piece of content that it either generates, edits, or passes along, and the end result at a destination, is that content is either 'trusted' if its cryptographic chain i…

If I'm like a high school/uni student, where this seems to really matter at the moment, you can still just have the LLM generate it and type it word for word in whatever text editor you're using, right? Less convenient but still probably easier than doing whatever work was necessary + still typing it all up. Chain of custody would show my keyboard really typed each stroke or w/e, but the underlying work is still generated

Re: How Claude marks AI-generated content

#436
post #303

Earlier quoted context omitted.

Perhaps. But it seems like your beef is not with the presence of watermarking, it's with what people will use that watermarking for. You're not directly harmed by that blog post being labeled as AI generated. In a hypothetical (but unfortunately likely) world where everything has passed through an AI's digestive system, nobody would care. In the meantime, it is true that this takes something away from you. But it's s…

I would push back on the fact that nobody would care. It is clear that platforms are increasing creating AI-generated as a category and it will only get more precise over time. Platforms that have built trust/aura over what type of content they host will resort to this more as they get more flooded with ai generated low effort content. I recently wrote a blog on this actually: https://decodingvibes.com/blog/aura-and-…

I'm going to push back on your pushback and say you're not disagreeing with me.

I was referring to a hypothetical future where everything has gone through AI to some degree, and for the purpose of the argument it has been modified enough to embed a watermark. Thus, the "detector" just always returns "true" in practice, and therefore nobody cares about what it says anymore.

I think what you call "aura" in your essay is already very much a thing. We use the adjective "artisanal", for example. Artisanal, bespoke, heirloom. hand-crafted, small batch, ... we have many ways of trying to convey "aura". (Heck, even things like "minority-owned" or "woman-owned" are attempts to imbue an aura.)

But more to the point of text, I'm also fascinated by the evolution of taste in this area. People's ability to detect AI slop improves, triggers revulsion to things they were previously fine with, and then sometimes they'll go back to being neutral once they start using AI more heavily themselves. Some AI writing is noticeably better quality than the human written version, yet it still smells of slop, and I at least am sometimes torn about which I prefer. Then there's the trend of AI writing getting better, substantially better, and sometimes at the same time worse in certain ways. Not to mention students, many of whom are now much more accustomed to AI writing than human writing -- it's their mental model of what an essay looks like, especially when all of their own essays uncoincidentally read like AI writing. And then there's all the nuance of the sometimes subtle differences between a wholly AI-generated piece of writing vs a heavily human-directed piece of AI writing vs an AI-edited piece of human writing vs a translated piece of human writing. It's fascinating and scary.

I too am curious about the final question in your essay: "will such categorization lead to the development of a potential human premium?"

Re: How Claude marks AI-generated content

#437
post #369

Earlier quoted context omitted.

I think we need a chain of custody system for content, but that would require browsers, software, websites, operating systems, phones, camera manufacturers, etc to all get on board. But each intermediary or source (optionally) cryptographicaly signs a piece of content that it either generates, edits, or passes along, and the end result at a destination, is that content is either 'trusted' if its cryptographic chain i…

If I'm like a high school/uni student, where this seems to really matter at the moment, you can still just have the LLM generate it and type it word for word in whatever text editor you're using, right? Less convenient but still probably easier than doing whatever work was necessary + still typing it all up. Chain of custody would show my keyboard really typed each stroke or w/e, but the underlying work is still gene…

The words themselves are the symbols that are used to calculate signature, copying word for word will still reveal provence.

Re: How Claude marks AI-generated content

#438
This is becoming serious anthropic hasn't even drop the anti detected pattern or mechanism this platform https://founderstoday.org/claude-watermark-remover is claiming to detect it using their own internal pattern but i tried it though it result is quite okay i can't even tell if it worked or not because the text are the same am wondering what they removed

Re: How Claude marks AI-generated content

#439

We need to just stop pretending we can reliably tell if plain text is written by an LLM. It’s just not a reasonable ask.

An approach like C2PA is the only realistic path forward. If the trajectory we're on continues, it's probably safe to assume nearly all content will be AI generated. We need realistic ways to prove content is human generated, and without true authentication (someone willing to corroborate they created the content, and they can certify it), the whole endeavor is pointless. While private human-verifiable content will still cease to exist, at least in this way we can avoid moving into an information dark-age
Post reply on HN