Live data from Hacker News

OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

ykilcher.com

21–30 of 64 posts

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#23
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…

Although FOSS is great, extreme wealth inequality has to be fixed by the government and not by open source developers.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#24
post #12

Earlier quoted context omitted.

Do you have any evidence that this is GPT-3.5 level, or are you just repeating what they said? We have an abundance of claimed capabilities already; that's not what's lacking.

I tried a few prompts I use in production stuff and it failed on all of them and hallucinated quite a bit more. All of these models are optimized for the gimmicky chatbot stuff that seems impressive to a casual user, but not for comparable capabilities to GPT-3.5. I wish what the parent said was true because it would save me money!

Which open model comes closest to GPT-3.5 in your production workload, if you don't mind me asking?

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#25

Earlier quoted context omitted.

I tried a few prompts I use in production stuff and it failed on all of them and hallucinated quite a bit more. All of these models are optimized for the gimmicky chatbot stuff that seems impressive to a casual user, but not for comparable capabilities to GPT-3.5. I wish what the parent said was true because it would save me money!

Which open model comes closest to GPT-3.5 in your production workload, if you don't mind me asking?

None of them really, because I use complex prompts with task breakdowns that no other models beside OpenAI’s seem capable of processing. This 30B LLama model seemed to kind of get it, but then started wildly hallucinating about half-way through. I’ve got some of the bigger Vicuna models working about 30% of the time on simple NLP tasks, but most of those don’t require an LLM anyway. They might perform better if you fine-tune them for whatever particular job, but that kind of defeats the purpose. The advantage of LLMs is supposed to be their generalized capabilities.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#26
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

I can't wait for an ad companies to force you to watch a 10 second video ad before it gives you a result for your query.

It's only a matter of time before these AI companies start pairing up with ad companies(if they already haven't). Google could easily put ad videos every 10 queries or something. You already see these limited free tokens/credits/querie on AI art sites.

How long until they put some ads in-between queries?

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#27
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…

Won't open-source AIs have their code stolen by the private AIs? There's no one stopping open-source AIs being used within private AIs.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#28
post #12
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

Do you have any evidence that this is GPT-3.5 level, or are you just repeating what they said? We have an abundance of claimed capabilities already; that's not what's lacking.

Section E of the paper we are "discussing" here.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#29
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

I can't wait for an ad companies to force you to watch a 10 second video ad before it gives you a result for your query. It's only a matter of time before these AI companies start pairing up with ad companies(if they already haven't). Google could easily put ad videos every 10 queries or something. You already see these limited free tokens/credits/querie on AI art sites. How long until they put some ads in-between qu…

At least showing an ad here and there is fair as you're using something for free and it's not trying to manipulate you, only show you an ad. It would be much worse if they actually biased the answers to talk about stuff they are paid to endorse...
Post reply on HN