(My apologies if this was already asked - this thread is huge and Find-In-Page-ing for variations of "pre-train", "pretrain", and "train" turned up nothing about this. If this was already asked I'd super-appreciate a pointer to the discussion :) ) Genuine question: How is it possible for OpenAI to NOT successfully pre-train a model? I understand it's very difficult, but they've already successfully done this and they…
I’m not sure what ‘successfully’ means in this context. If it means training a model that is noticeably better than previous models, it’s not hard to see how that is challenging.
I can totally see how they're able to pre-train models no problem, but are having trouble with the "noticeably better" part.
Thanks!