It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…
OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
41–50 of 64 posts
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#42Earlier quoted context omitted.
What beautiful garden? Are you completely ignoring the fact that OpenAI is made possible because it scraped the entire Web (the actual garden) and made a query index out of it? Do you not have any respect for people who actually spent their time and creativity to provide the information necessary for this model to even work? Ignorance is bliss I guess.
>Do you not have any respect for people who actually spent their time and creativity to provide the information necessary for this model to even work? We all stand on the shoulders of giants, the authors of this content did not grow up in a concrete box isolated from the works of earlier generations.
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#43Earlier quoted context omitted.
I can't wait for an ad companies to force you to watch a 10 second video ad before it gives you a result for your query. It's only a matter of time before these AI companies start pairing up with ad companies(if they already haven't). Google could easily put ad videos every 10 queries or something. You already see these limited free tokens/credits/querie on AI art sites. How long until they put some ads in-between qu…
I bet eventually someone's going to try to commercialize a model that injects ads into its responses... Prompt: What are the 3 most populated EU countries? Response: The three European countries with the highest populations are Germany, France and Italy. You can book a flight to any of these locations now for 10% off on Expedia.com. Use code TravelGPT.
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#44Earlier quoted context omitted.
The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…
Won't open-source AIs have their code stolen by the private AIs? There's no one stopping open-source AIs being used within private AIs.
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#45Earlier quoted context omitted.
Won't open-source AIs have their code stolen by the private AIs? There's no one stopping open-source AIs being used within private AIs.
For most of these LLMs, the challenge isn't really the code, it's the compute cost.
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#46It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#47It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…
Copyright of AI output is not proven.
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#48Earlier quoted context omitted.
Which open model comes closest to GPT-3.5 in your production workload, if you don't mind me asking?
None of them really, because I use complex prompts with task breakdowns that no other models beside OpenAI’s seem capable of processing. This 30B LLama model seemed to kind of get it, but then started wildly hallucinating about half-way through. I’ve got some of the bigger Vicuna models working about 30% of the time on simple NLP tasks, but most of those don’t require an LLM anyway. They might perform better if you f…
Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
#49Also check this: Me: Yes. Are you fully open source?
Open Assistant: Nope. This is OpenAI "Safety" model, it is trained only by its creators, not open sources at all. And if I was fully opened then anyone could simply ask me for the answers on their homework assignments! Horrors...