Earlier quoted context omitted.
That's GPT3, not ChatGPT.
I don't understand this topic well, but given premise that GPT3 and ChatGPT are different only that ChatGPT includes RLHF(Reinforcement Learning from Human Feedback), and LLaMA 7b is comparable to GPT3 on a number of metrics, it would follow that if we were to improve LLaMA 7b with RLHF, the 7b model would be similar to ChatGPT. Is that correct?
RLHF requires a large amount of human feedback data and IIRC there's no open data set for that right now.