Live data from Hacker News

Curious about the training data of OpenAI's new GPT-OSS models? I was too

twitter.com

61–62 of 62 posts

Re: Curious about the training data of OpenAI's new GPT-OSS models? I was too

#61
post #38
post #28

OP seems to have run a programming language detector on the generated texts, and made a graph of programming language frecuencies: https://pbs.twimg.com/media/Gx2kvNxXEAAkBO0.jpg?name=orig As a result, OP seems to think the model was trained on a lot of Perl: https://xcancel.com/jxmnop/status/1953899440315527273#m LOL! I think these results speak more to the flexibility of Perl than any actual insight on the training…

I don't understand why Perl, R, and AppleScript rank so much higher than their observed use.

It seems to be an error with the classifier. Sorry everyone. I probably shouldn't have posted that graph; I knew it was buggy, I just thought that the Perl part might be interesting to people.

Here's a link to the model if you want to dive deeper: https://huggingface.co/philomath-1209/programming-language-i...

Re: Curious about the training data of OpenAI's new GPT-OSS models? I was too

#62

"this thing is clearly trained via RL to think and solve tasks for specific reasoning benchmarks. nothing else." Has the train already reached the end of the line?

If you think something like "They have to train their models on benchmarks to make it look like there's progress, while in reality it's a dead end," you are missing a few things.

It's an open model, everyone can bench it on everything not only on specific benchmarks. Training on specific reasoning benchmarks is a conjecture.

Post reply on HN