The article says "now in the wild after being leaked" but then it says "the data is impossible to retrieve as it is now stored on the servers belonging to OpenAI." So did the source code leak out of OpenAI into the wild, or are they saying that OpenAI itself is "the wild"? As far as I see from the article, it's not accessible to the general public.
If ChatGPT is trained on the data and ChatGPT is accessible to the general public, then the data may as well be accessible to the general public
I assume the chat logs are instead training a reward model, which itself is then used as the reward function during RLHF training.