Live data from Hacker News

OpenAI won't watermark ChatGPT text because its users could get caught

theverge.com

51–60 of 84 posts

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#51
post #46
post #33

Earlier quoted context omitted.

> Force all companies training LLMs to add some method of watermarking with a mean error rate below a set value. How do you watermark plain text?

"I seem to do fine for a stretch, but at the of the sentence I say the wrong cranberry."

So deliberately bork the output in such a way that users lose all confidence in the product?

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#52
If I understand the regulation correctly, they will have to add it in order to comply with the AI act:

> In addition, providers will have to design systems in a way that synthetic audio, video, text and images content is marked in a machine-readable format, and detectable as artificially generated or manipulated.

https://ec.europa.eu/commission/presscorner/detail/en/ip_24_...

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#53
Today, a colleague committed ten lines of code outputted by ChatGPT, without testing it, and the systems broke. Any chance watermarking would help with this? Just kidding (about the watermarking).

The AI cat is already out of the bag. I'm for regulation when it comes to AI direst threats, but students using ChatGPT to cheat is a problem that can be solved with live, supervised exams, the kind I had at school in the nineties...we couldn't even use a pocket calculator.

If anything, I'm torn that today's AI systems aren't good enough to do the really serious words. Mark my words, we will have self-aware murder robots before we have AI systems able to write quantum-simulation software.

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#54
post #33
post #13

What we're seeing is a new profitable industry, with a tendency towards natural monopoly, refusing to act in the public good because its incentives conflict with it. At stake is the entire corpus of 21st century media, at risk of being drowned out in a tsunami of statistically unidentifiable spam, which has already begun and will only get worse, until nobody will even acknowledge the web post 2022 as worth reading or…

> Force all companies training LLMs to add some method of watermarking with a mean error rate below a set value. How do you watermark plain text?

This is a fairly obvious initial question which I assume nearly everyone who doesn't already have a rough answer in mind would ask, so I'm happy to report that it's fortunately quite clearly addressed in TFA, and in fact makes up a significant part of the (not very long) piece.

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#55
post #33
post #13

What we're seeing is a new profitable industry, with a tendency towards natural monopoly, refusing to act in the public good because its incentives conflict with it. At stake is the entire corpus of 21st century media, at risk of being drowned out in a tsunami of statistically unidentifiable spam, which has already begun and will only get worse, until nobody will even acknowledge the web post 2022 as worth reading or…

> Force all companies training LLMs to add some method of watermarking with a mean error rate below a set value. How do you watermark plain text?

[deleted]

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#56
post #7

On a personal level, I'm happy about this. Having all personal info trackable is a tiring aspect of modern life? Remember when social media didn't zero out the exif on photos and anyone could easily grab a map to your family's home? On a society level? I'm stil ok with this. Watermarks are trivially removed by the motivated and unpleasant.

Its hard to remove this watermark, because the "watermark" is adjusting the probabilities of what tokens are generated, rather than just slapping "generated by chatgpt" over the top. You'd have to actually rewrite the text to remove it

so the fact that it's watermarked literally means it was tampered with to produce some form of predetermined output. No other way to do it.

And because they are entirely uncontrollable they have to add a lot of checks in the prompts. but we all know that hasn't worked. A series of manually built if statements is akin to an expert system. The first thing I learned about those decades ago was that they were mostly a failed experiment

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#57

Earlier quoted context omitted.

GPT-2 was never too dangerous to release, that's made up. OpenAI were trying to set a precedent to delay releases in anticipation of more dangerous models. This was and remains a good thing.

Not sure about remains. The Llama is out of the bag with that one! OpenAI models aren't the best all-round anymore - there are plenty of faster models to choose from. Claude seems faster. Groq is stupid fast in serving Llama. gpt4 still has the smarts edge, but it isn't much of an edge any longer. OpenAI is iPhone, it is 2019 and people are realizing the Androids are better value and just as good

I don't exactly see Llama as a counterexample. Not much wisdom coming from Facebook in this matter.

Like, yay we got Sydney back! Everybody gets their own death threats. It's an open-weights wonderland. Why people think it's a good idea to reinstantiate the BPD AI I am afraid I will never understand.

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#58
post #48

Earlier quoted context omitted.

Why exactly is it insane ? To reliably differentiate (let's assume it's possible for the sake of argument) between "you made this" and "you didn't make this" or at least "a human made this" seems to carry mostly (if not only) benefits.

the problem is your parenthetical - it's not possible, so attempting to do so isn't actually really possible. what's worse than a watermark? one that doesn't actually work.

Open AI literally said they have a semi-resilient method with 99.9% accuracy. It will become full-resilient for practical purposes if all LLMs implement something similar.

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#59
post #51
post #46

Earlier quoted context omitted.

"I seem to do fine for a stretch, but at the of the sentence I say the wrong cranberry."

So deliberately bork the output in such a way that users lose all confidence in the product?

You should read the original article... OpenAI say it doesn't downgrade quality.

Re: OpenAI won't watermark ChatGPT text because its users could get caught

#60
post #53

Today, a colleague committed ten lines of code outputted by ChatGPT, without testing it, and the systems broke. Any chance watermarking would help with this? Just kidding (about the watermarking). The AI cat is already out of the bag. I'm for regulation when it comes to AI direst threats, but students using ChatGPT to cheat is a problem that can be solved with live, supervised exams, the kind I had at school in the n…

> Today, a colleague committed ten lines of code outputted by ChatGPT, without testing it, and the systems broke

Do you work at Crowdstrike?

But seriously, I can only imagine how bad their code would have been without ChatGPT.

Post reply on HN