Live data from Hacker News

OPT: Open Pre-trained Transformer Language Models

arxiv.org

1–10 of 242 posts

Re: OPT: Open Pre-trained Transformer Language Models

#2
"We are releasing all of our models between 125M and 30B parameters, and will provide full research access to OPT-175B upon request. Access will be granted to academic researchers; those affiliated with organizations in government, civil society, and academia; and those in industry research laboratories."

GPT-3 Davinci ("the" GPT-3) is 175B.

The repository will be open "First thing in AM" (https://twitter.com/stephenroller/status/1521302841276645376):

https://github.com/facebookresearch/metaseq/

Re: OPT: Open Pre-trained Transformer Language Models

#5

"We are releasing all of our models between 125M and 30B parameters, and will provide full research access to OPT-175B upon request. Access will be granted to academic researchers; those affiliated with organizations in government, civil society, and academia; and those in industry research laboratories." GPT-3 Davinci ("the" GPT-3) is 175B. The repository will be open "First thing in AM" ( https://twitter.com/stephe…

I don't like "available on request". I just want to download it and see if I can get it to run and mess around with it a bit. Why do I have to request anything? And I'm not an academic or researcher, so will they accept my random request?

I'm also curious to know what the minimum requirements are to get this to run in inference mode.

Re: OPT: Open Pre-trained Transformer Language Models

#6

Does this make Meta AI more “open” than OpenAI? Oh, the irony.

"Open" in OpenAI is like countries with "Democratic" in their name e.g. Democratic People's Republic of Korea

https://petervojtek.github.io/diy/2015/05/19/countries-with-...

Re: OPT: Open Pre-trained Transformer Language Models

#8

"We are releasing all of our models between 125M and 30B parameters, and will provide full research access to OPT-175B upon request. Access will be granted to academic researchers; those affiliated with organizations in government, civil society, and academia; and those in industry research laboratories." GPT-3 Davinci ("the" GPT-3) is 175B. The repository will be open "First thing in AM" ( https://twitter.com/stephe…

I don't like "available on request". I just want to download it and see if I can get it to run and mess around with it a bit. Why do I have to request anything? And I'm not an academic or researcher, so will they accept my random request? I'm also curious to know what the minimum requirements are to get this to run in inference mode.

My bet it's probably a filter, trying to prevent create a even more realistic farmbots in social media, as they are already bad as they are now.

Re: OPT: Open Pre-trained Transformer Language Models

#9
post #8

Earlier quoted context omitted.

I don't like "available on request". I just want to download it and see if I can get it to run and mess around with it a bit. Why do I have to request anything? And I'm not an academic or researcher, so will they accept my random request? I'm also curious to know what the minimum requirements are to get this to run in inference mode.

My bet it's probably a filter, trying to prevent create a even more realistic farmbots in social media, as they are already bad as they are now.

But they'll consider requests from government and industry.. both greater threats in the information war than any private individual.

Re: OPT: Open Pre-trained Transformer Language Models

#10

"We are releasing all of our models between 125M and 30B parameters, and will provide full research access to OPT-175B upon request. Access will be granted to academic researchers; those affiliated with organizations in government, civil society, and academia; and those in industry research laboratories." GPT-3 Davinci ("the" GPT-3) is 175B. The repository will be open "First thing in AM" ( https://twitter.com/stephe…

I don't like "available on request". I just want to download it and see if I can get it to run and mess around with it a bit. Why do I have to request anything? And I'm not an academic or researcher, so will they accept my random request? I'm also curious to know what the minimum requirements are to get this to run in inference mode.

Couple of random ideas:

- They are concerned about the usage of the largest model, so want to vet people

- The 175B parameter model is so large that it doesn't play nice with GitHub or something along those lines

Post reply on HN