Live data from Hacker News

OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

ykilcher.com

31–40 of 64 posts

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#31
Does anyone have any tips for how to spin up services that can efficient peform inference with the HuggingFace weights of models like this.

I would love to switch to something like this over OpenAI's GPT3.5 Turbo, but this weekend I'm struggling to get reasonable inference speed on reasonably priced machines.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#32
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…

[deleted]

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#33
post #4

This makes Turbo GPT 3.5 level AI free, private and finetuneable. OpenAI exclusivity shrinks now to GPT-4. That's why I don't think they will be able to keep a large market share in LLMs, any level of AI is going to get open and free soon. SOTA models are also easy to distill by API, very hard to defend against using chat logs as training data for other models. Once we all got one running, maybe in the OS, maybe in t…

I can't wait for an ad companies to force you to watch a 10 second video ad before it gives you a result for your query. It's only a matter of time before these AI companies start pairing up with ad companies(if they already haven't). Google could easily put ad videos every 10 queries or something. You already see these limited free tokens/credits/querie on AI art sites. How long until they put some ads in-between qu…

I bet eventually someone's going to try to commercialize a model that injects ads into its responses...

Prompt: What are the 3 most populated EU countries?

Response: The three European countries with the highest populations are Germany, France and Italy. You can book a flight to any of these locations now for 10% off on Expedia.com. Use code TravelGPT.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#34

Earlier quoted context omitted.

The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…

Won't open-source AIs have their code stolen by the private AIs? There's no one stopping open-source AIs being used within private AIs.

Yes. And then the private AI will quickly lag behind as the open-source AI is continuously updated.

Even if the private AI owner made some unique discovery which gives them an advantage, its very likely to only be unique for a short while (see: some of the world’s major discoveries simultaneously found by different people. I’m sure there would be more if not for word-of-mouth)

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#35
Congrats to all the people behind this!

To the naysayers that point out that OpenAssistant is not as good as ChatGPT: all of LLM development is still in the early days; the work being done by teams like OA is helping to shape the future development of the entire field by creating an actually-open-source foundation. The results of this work will only get better with time.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#36
post #17
post #5

Earlier quoted context omitted.

What beautiful garden? Are you completely ignoring the fact that OpenAI is made possible because it scraped the entire Web (the actual garden) and made a query index out of it? Do you not have any respect for people who actually spent their time and creativity to provide the information necessary for this model to even work? Ignorance is bliss I guess.

Out of curiosity, do you use ad blocking software?

No. See the web how your typical user does

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#37

What is the token limit? The 2k limit on llama is *very limiting on the number of things it can do.

One of the main models here is LLama so it’s limited to 2K tokens. Not sure about the Pythia one.

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#38
post #23

Earlier quoted context omitted.

The problem is that for-profit businesses like OpenAI have more money and compute than even millions of volunteers. I definitely believe we'll get an open GPT-4 eventually, but by then OpenAI will have GPT-5, and so on. It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain…

Although FOSS is great, extreme wealth inequality has to be fixed by the government and not by open source developers.

[deleted]

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#39
I had quite some fun asking questions and finding the limits of it's (current) knowledge. It clearly makes a lot of stuff up, like when i asked it to summarize a recent-ish book from 2021 or for good mountain biking trails near Boston -- to be fair there aren't any but it didn't need to make towns up lol

With more RIFL it will only get better. Nice progress!!

Re: OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]

#40

It's awesome that the OpenAssistant project made it this far with a lot of crowed-sourced input. Congrats to the whole team that works really hard trying to create a truly open LLM. One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purp…

> due to the OpenAI license used to generate the data.

What makes you think OpenAI responses are copyrighted in any way?

Post reply on HN