I'm confused why people expect this stuff to be free? I'm surprised OpenAI was so open about their research so far. I don't blame them at all for not publishing the information. This stuff costs real money.
We don't expect it to be free -- please read the article. That's not the issue at all. It's like if you subscribe to a product that you need to do your job, and one day the company tells you that the product is going away in three days and that you need to switch to a different product (that isn't at all the same for your use case).
OpenAI’s policies hinder reproducible research on language models
61–70 of 394 posts
Re: OpenAI’s policies hinder reproducible research on language models
#62IMO established companies (Meta, Google, etc) had their researchers publish papers as a competitive benefit or way to attract talent from academia (a researcher wouldn't want to stop publishing). Companies didn't see an issue with doing that because those papers were not "giving away" the core of the company, for example, Facebook's DeepFace paper from 2014 couldn't hurt its ad business. OpenAI on the other hand will…
Re: OpenAI’s policies hinder reproducible research on language models
#63Earlier quoted context omitted.
You say: OpenAI really seems to believe that not publishing details will save them from competition. Then say: The only competitive advantage that OpenAI has here is a headstart of 6-12 months It's almost as if they want to keep this advantage, huh? Blows my mind how business illiterate some HN commenters are.
A headstart doesn't matter unless you can keep it. The point is that there are many mort smart people and resources outside of OpenAI and there are inside of OpenAI. If they focus their efforts, they will easily catch up. A headstart is not a competitive moat like network effects are. Go and try to raise money for your startup from a VC and tell them "well, everyone is doing the same as us, but we started 6 months ea…
Re: OpenAI’s policies hinder reproducible research on language models
#64Open AI has been doing sketchyish things long before Chat GPT, and I think it's something people are eventually going to notice more and more (then again people were swearing that Musk walked on water for waaaaaay too long given his actions so fuck if I know). They're 100% marketing FIRST. I don't think they'll outright lie, but they will absolutely screw with their data in such a way to make it look waaay more impre…
Re: OpenAI’s policies hinder reproducible research on language models
#65Since OpenAI didn't release the parameter count of GPT-4, I've been wondering/doubting if it is really much bigger than GPT-3. The release of GPT-3.5 has shown that they've found ways of drastically cutting down compute costs (an order of magnitude) while maintaining or even improving the quality of the model's outputs. Perhaps the reason that they didn't release the specifics of GPT-4 might be in part due to them wa…
Re: OpenAI’s policies hinder reproducible research on language models
#66But stuff got real. They discovered a path to super-human cognition that scales directly with money and computer chips. Now these companies are closing their public academic work, looking for partnerships with companies like nvidia, and firing large swaths of employees.
Re: OpenAI’s policies hinder reproducible research on language models
#67I understand any individual's company anti-competitive measures. OpenAI looks at Google the same way Apple looked at IBM in the 80s. What I'm worried about is a lot of the talk about guarding models, public safety and misuse of models will end up leading every big company to pull public access of their APIs. We might look at 2022-2023 as a brief golden age when regular people could use stuff like GPT-4 before it was…
Re: OpenAI’s policies hinder reproducible research on language models
#68What's most surprising to me is that OpenAI really seems to believe that not publishing details will save them from competition. Everyone knows how these models work, and while I'm sure there is a bunch of "secret sauce" that OpenAI has built for training and fine-tuning, it's ridiculous to believe that the research community and competitors like Google and Facebook can't figure out the same. They just haven't really…
Au contraire, no one knows how large GPT-4 is, which is the single best predictor of performance (for a model trained to convergence). The GPT-4 paper spent much of its time writing about this — they did some small scale experiments with 1/1000th the compute, then picked a loss level they wanted and trained GPT-4 till it got it. Neither the exact loss level nor the number of parameters are revealed by the paper. Unfo…
As you say, there is absolutely no way for us to find out.
Re: OpenAI’s policies hinder reproducible research on language models
#69Re: OpenAI’s policies hinder reproducible research on language models
#70I understand any individual's company anti-competitive measures. OpenAI looks at Google the same way Apple looked at IBM in the 80s. What I'm worried about is a lot of the talk about guarding models, public safety and misuse of models will end up leading every big company to pull public access of their APIs. We might look at 2022-2023 as a brief golden age when regular people could use stuff like GPT-4 before it was…
The stuff about models potentially being misused is just their public justification to look like the good guys. They're not going to withhold their technology because they don't want it to be misused, they're withholding it because they want control over who misuses it. Of course, it won't be called "misuse" when the right parties are doing it.