Earlier quoted context omitted.
Maybe the authors of the text should also have a stake in GPT-3. After all, OpenAI didn't write the corpus. Google benefits from the web, and Facebook from the real-life social networks and their activities that it replicates online (messages, meetings, news, etc). We are all the source of the training data. Why should we be at the whims of these derivative product companies?
Fun fact, OpenAI did not collect the data itself - they mainly used data from Common Crawl (in addition to a couple other datasets), which is compiled by a non profit that shares the dataset for free. So perhaps the license of such datasets can encourage free sharing of research outcomes. https://commoncrawl.org/
I think ml models should be public unless the data itself isn't.