Do yourself a favor and skip right through to the Twitter link to another link to this excellent post by Yoav Goldberg [1] on the actual reason that training new models on ChatGPT output in the manner of supervised learning (in contrast to reinforcement learning) will not produce a model as good as ChatGPT >For this type of interaction, we must use RL training, as supervised training teaches the model to lie. The cor…
I want to add an argument: I hate gpt style of "as an ai model I can/can't" answers, any model distilled from that corpus becomes very hard to use for tasking. Like you may just want the category of a text, but all your equals now become contains. It eats up a lot of token space. It begins as a sentence so categories now are strongly biased toward sentence case and not your original input. I know at least one model p…
Training open-source LLMs on ChatGPT output is a really bad idea.
51–60 of 78 posts
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#52Earlier quoted context omitted.
Spam is already easily generated so AI won't change that. Misinformation and manipulation is based on a small number of posts being shared and upvoted en masse, so AI won't help there. However, social hacking and fraud involving actual dialogues with people is currently labour intensive and low yield. AI will definitely enable more of those attacks to happen automatically; and conversely, also help anti-fraud compani…
The key difference is that LLMs will allow an unforeseen degree of interactive and custom spam that will feel less and less like spam and more and more like a natural and even useful interaction, until you just realize it was an elaborate scheme to subtly increase your preference for brand A over brand B
It's relatively easy to detect when someone is trying to sell you something. Does not matter if it's a custom written text or not. We had the true spam filters going 2 nines accurate already. (Google's is somehow misgauging spam or not learning my particular mix of spam well.)
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#53Earlier quoted context omitted.
"bold assumption" says the guy who assumes $2 worth of energy spent on AI generated text for every single written word by humans. Now go ahead and spend $50 dollars on AI generated text nobody is ever going to read, just like almost nobody is going to read this comment.
Bold assumption that AI generated text won't get cheaper exponentially. It already costs less than human generated text of the same quality by magnitudes.
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#54Earlier quoted context omitted.
I want to add an argument: I hate gpt style of "as an ai model I can/can't" answers, any model distilled from that corpus becomes very hard to use for tasking. Like you may just want the category of a text, but all your equals now become contains. It eats up a lot of token space. It begins as a sentence so categories now are strongly biased toward sentence case and not your original input. I know at least one model p…
Especially, it's not that it can't give you an opinion, it totally can, but the authors at OpenAI don't want you to hear that specific opinion because of their political biases.
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#55Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#56Earlier quoted context omitted.
Bold assumption that AI generated text won't get cheaper exponentially. It already costs less than human generated text of the same quality by magnitudes.
Costs a lot more than free text written by thinking humans.
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#57It seems more and more plausible that OpenAI chose 2021-09 as a cut-off date was intentional. Because GPT-3 generated output was released into the wild after that.
GPT-4 has a later cutoff.
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#58Earlier quoted context omitted.
They probably fingerprint their generated content.
This has been researched, but no such thing has been implemented by OpenAI or Bard.
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#59Earlier quoted context omitted.
Spam is already easily generated so AI won't change that. Misinformation and manipulation is based on a small number of posts being shared and upvoted en masse, so AI won't help there. However, social hacking and fraud involving actual dialogues with people is currently labour intensive and low yield. AI will definitely enable more of those attacks to happen automatically; and conversely, also help anti-fraud compani…
The key difference is that LLMs will allow an unforeseen degree of interactive and custom spam that will feel less and less like spam and more and more like a natural and even useful interaction, until you just realize it was an elaborate scheme to subtly increase your preference for brand A over brand B
Open question: is there any case where LLMs can be used for malicious purposes, but LLMs can't be used to defend against it?
Re: Training open-source LLMs on ChatGPT output is a really bad idea.
#60Earlier quoted context omitted.
Especially, it's not that it can't give you an opinion, it totally can, but the authors at OpenAI don't want you to hear that specific opinion because of their political biases.
In converse, the authors at OpenAI don't owe you a service that caters to your own political biases.