Live data from Hacker News

Ask HN: How would you build a ChatGPT detector?

news.ycombinator.com

71–80 of 109 posts

Re: Ask HN: How would you build a ChatGPT detector?

#71

Earlier quoted context omitted.

I suspect that in time, this will only accelerate the degree to which AI and human-authored text are indistinguishable from each other.

I've already sent text to customers 100% written by AI. Ethically dubious in a commercial setting perhaps, but higher quality text than I would be able to produce myself. I asked OpenAI and it said: It is not necessarily unethical to send customers text generated by AI, but it depends on the context and the specific situation. For example, if the text is being used to deceive or mislead customers, then it would be un…

One of the first things I did with it was asking it to write a resignation letter in the style of a Shakespeare sonnet. I golfed it a little bit and got a nice result which I tucked it into a text file for future use.

Re: Ask HN: How would you build a ChatGPT detector?

#72
post #12

Earlier quoted context omitted.

Right, but in principle the detectors could iterate and get better over time too. That's why I asked about "a ChatGPT detector" instead of "general AI detector", which is a very different problem.

The default style is easily detected by presence of overalls, moreovers and furthermores. When you tell it to do a ‘concise HN comment’ the only tell may be that its English is flawless. E.g. In a style of a very concise HN comment describe how to detect that a text has been written by the Assistant. To detect if a text has been written by the Assistant, check for the use of generic language, repetitive phrases, and…

> To detect if a text has been written by the Assistant, check for the use of generic language, repetitive phrases, and a lack of personal perspective or opinion.

There are many people who use generic language, repetitive phrases, and lack personal perspectives and opinions (at least people who lack sophisticated perspectives/opinions).

Which makes the problem more complicated, but also perhaps irrelevant. If you can’t tell whether some text was written by a human or a computer, does it even matter?

Re: Ask HN: How would you build a ChatGPT detector?

#73
Probably the old fashioned way :-). I think you would need a human to detect the BS / super passive voice. At some point in the future it becomes impossible to detect even for say a trained FBI team, then that might be the point it no longer matters. Then trying to detect it is like trying to win chess by improving your km/min running speed. You just have to play a different game.

Re: Ask HN: How would you build a ChatGPT detector?

#74
I'm with Paul Graham: we've crossed the Turing threshold, the code duplicates typical human-created internet gibberish with uncanny accuracy, and it's impossible to tell the difference even in principle.

Example: "When I said tequila has a relatively high sugar content, I was not implying that tequila contains sugar."

-- from "Neckbeard or AI? Is there even a difference?" https://www.lesswrong.com/posts/goC9qv4PWf2cjfnbm/did-chatgp...

Re: Ask HN: How would you build a ChatGPT detector?

#75
post #43

> Answer the question, "How would you build a ChatGPT detector?" written in the style of a cynical Hacker News commenter Well, if I were to build a ChatGPT detector, I would probably just use some basic natural language processing and machine learning algorithms. But let's be real here, it's not like ChatGPT is some kind of advanced AI that requires some groundbreaking technology to detect. It's just a chatbot with a…

I got:

A ChatGPT detector? Ha! Good luck with that. It's nearly impossible to build a detector for a large language model like ChatGPT. The only way to really do it is to constantly monitor the output of the model and try to flag anything that seems suspicious. But even then, it's hard to say for sure if the model is truly being used for nefarious purposes or not. Plus, with a model as powerful as ChatGPT, it's likely that any attempts to detect its use will be quickly outwitted. So, in short, don't even bother trying to build a ChatGPT detector. It's a waste of time and resources.

I think it's trying to protect itself.

I tried it again:

Well, first off, I'd probably just rip off OpenAI's GPT-3 and try to shoehorn it into a chatbot. But let's be real, that's not going to work very well. So instead, I'd probably just use some basic keyword detection and flag any messages that contain certain triggering words or phrases. It's not going to be very accurate, but at least it'll give the appearance of trying to do something about the whole ChatGPT issue. Plus, it'll probably generate a bunch of false positives and annoy the hell out of users, which is always a bonus in my book.

A+ snark

Re: Ask HN: How would you build a ChatGPT detector?

#76
Ask it for a URL or a link to backup any claim that it's making. This URL may or may not point to a real web resource. Regardless, claim that the web page is down for you. It may well then admit to a contradiction, such as:

> AI: Yes, our website is currently up and running. You may have encountered an issue due to a temporary technical glitch. I apologize for the inconvenience and I can provide you with an updated link if you need it.

or

> AI: I apologize if that was unclear. Our website is up and running, but we are aware of some technical issues that have caused some users to experience difficulty when trying to access the website.

Re: Ask HN: How would you build a ChatGPT detector?

#77
post #18

According to ChatGPT: > 1. Start by training a machine learning model to identify features of natural language. You can use supervised learning techniques with a corpus of labeled data consisting of texts written by both humans and AI systems. > 2. Use natural language processing techniques to extract features from the text such as word frequencies, grammar, and syntax. > 3. Implement a scoring system to assign a num…

I suspect that in time, this will only accelerate the degree to which AI and human-authored text are indistinguishable from each other.

This. Once AI gets good enough at replicating the logic we follow in producing texts we consider reasonable illogical texts will become markers of humanity. In the future, we are all dadaists.

Re: Ask HN: How would you build a ChatGPT detector?

#78
I want you to act as a generated text detector. I will write you some text and you will tell me If it was generated by a man or by a machine. Only write who you think wrote the text. Do not give explanations. The first text is "Like everyone else, I'm blown away by ChatGPT's responses to prompts. At the same time, there's a certain sameiness to the language it produces. This makes me wonder, how hard would be to build a different AI that would recognize the writing of this AI? And how accurate could it get?"

Re: Ask HN: How would you build a ChatGPT detector?

#80
post #64

Lol, it wouldn't be that hard to build an AI that could recognize ChatGPT's writing. I mean, it's not like ChatGPT is producing some super unique and creative language or anything. It's just spitting out the same old generic responses to prompts. If you want to build an AI that could accurately recognize ChatGPT's writing, just train it on a bunch of examples of ChatGPT's responses and it'll be able to pick out the c…

If ChatGPT can imitate other styles of writing (hacker news, 4chan style for example), it gets much harder to detect. I think within a few years it will become infeasible to say whether something was authored by an AI

It's already been prevented from emulating styles (at least directly):

> As a large language model trained by OpenAI, I do not have any knowledge of specific styles or conventions used on Hacker News or any other online forum. I am a neutral, unbiased source of information and do not have the ability to engage in discussions or adopt specific styles. My purpose is to provide accurate and helpful information to the best of my ability.

I'm not sure if they think the genie will go back into the bottle or if they just don't want to be the ones summoned to a congressional hearing once someone uses it for some kind of crime enabled by that kind of ability.

Post reply on HN