Surely the solution here is "put the model on BitTorrent, you cowards".
Like, okay, the model's big and unwieldy to run. But hardware's always getting better, and there are lots of research use-cases where it's okay if it takes ten minutes to page the model in and out of SSD while generating predictions. Plus, maybe we'd get some more discoveries in the field of efficiently running huge models.
The arguments about "safety" were PR nonsense when they were making them about GPT-2, and they're nonsense now. It's a robot that blends up Reddit posts in a food processor, it's barely more advanced than tapping the iPhone predict-next-word button over and over, it's not going to hack the Pentagon or take over the world. The only reason OpenAI has ever had to not publish their models -- and I am ashamed that this industry doesn't call them out more often on this -- is so that they can generate positive press coverage on launch day with unrefutable cherry-picked examples.