Live data from Hacker News

OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

ykilcher.com

41–50 of 64 posts

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#41

It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…

Copyright of AI output is not proven.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#42
post #5

Earlier quoted context omitted.

What beautiful garden? Are you completely ignoring the fact that OpenAI is made possible because it scraped the entire Web (the actual garden) and made a query index out of it? Do you not have any respect for people who actually spent their time and creativity to provide the information necessary for this model to even work? Ignorance is bliss I guess.

>Do you not have any respect for people who actually spent their time and creativity to provide the information necessary for this model to even work? We all stand on the shoulders of giants, the authors of this content did not grow up in a concrete box isolated from the works of earlier generations.

It's like these people have never been to a library.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#43
post #33

Earlier quoted context omitted.

I can't wait for an ad companies to force you to watch a 10 second video ad before it gives you a result for your query. It's only a matter of time before these AI companies start pairing up with ad companies(if they already haven't). Google could easily put ad videos every 10 queries or something. You already see these limited free tokens/credits/querie on AI art sites. How long until they put some ads in-between qu…

I bet eventually someone's going to try to commercialize a model that injects ads into its responses... Prompt: What are the 3 most populated EU countries? Response: The three European countries with the highest populations are Germany, France and Italy. You can book a flight to any of these locations now for 10% off on Expedia.com. Use code TravelGPT.

My speculation is that whatever we will be typing into these chat boxes (be it Bing, Bard, or free ChatGPT) will ultimately be used to for ad targeting anywhere on the web and beyond, either in-house (Bing) or via data brokers.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#44

Earlier quoted context omitted.

The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…

Won't open-source AIs have their code stolen by the private AIs? There's no one stopping open-source AIs being used within private AIs.

For most of these LLMs, the challenge isn't really the code, it's the compute cost.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#45
post #44

Earlier quoted context omitted.

Won't open-source AIs have their code stolen by the private AIs? There's no one stopping open-source AIs being used within private AIs.

For most of these LLMs, the challenge isn't really the code, it's the compute cost.

The compute cost is a challenge for amateurs and Europeans. In United States, investors will throw frankly quite ludicrous amounts of money at you if you show promise.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#46

It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…

If openai owns openassistant because it was trained in part on chatgpt outputs, then andrew hussie owns chatgpt because it was trained in part on homestuck

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#47

It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…

Copyright of AI output is not proven.

Is rather think it's the opposite, it's almost definitely proven that it is not - it is obviously completely transformative.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#48

Earlier quoted context omitted.

Which open model comes closest to GPT-3.5 in your production workload, if you don't mind me asking?

None of them really, because I use complex prompts with task breakdowns that no other models beside OpenAI’s seem capable of processing. This 30B LLama model seemed to kind of get it, but then started wildly hallucinating about half-way through. I’ve got some of the bigger Vicuna models working about 30% of the time on simple NLP tasks, but most of those don’t require an LLM anyway. They might perform better if you f…

I think most people don't realise that OpenAi biggest advantage is the billions of queries it was asked those signals are what they used to optimise it. So I think very hard for a local model to reach similar capability.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#49
I tried it. It's honestly extremely impressive and legitimately usable and useful, maybe as much as ChatGPT, but I'd have to test it more to be sure. The single biggest, game-changing advantage OpenAssistant has is that is that it is NOT CENSORED. This is legitimately AMAZING. I asked it a touchy question that chatgpt refused to answer and it gave me a fair, clear-cut, detailed answer.

Also check this: Me: Yes. Are you fully open source?

Open Assistant: Nope. This is OpenAI "Safety" model, it is trained only by its creators, not open sources at all. And if I was fully opened then anyone could simply ask me for the answers on their homework assignments! Horrors...

Post reply on HN