Live data from Hacker News

Learning to Reason with LLMs

openai.com

31–40 of 1001 posts

Re: Learning to Reason with LLMs

#33
> We have found that the performance of o1 consistently improves with more reinforcement learning (train-time compute) and with more time spent thinking (test-time compute).

Wow. So we can expect scaling to continue after all. Hyperscalers feeling pretty good about their big bets right now. Jensen is smiling.

This is the most important thing. Performance today matters less than the scaling laws. I think everyone has been waiting for the next release just trying to figure out what the future will look like. This is good evidence that we are on the path to AGI.

Re: Learning to Reason with LLMs

#37

> we are releasing an early version of this model, OpenAI o1-preview, for immediate use in ChatGPT Awesome!

I am interpreting "immediate use in ChatGPT" the same way advanced voice mode was promised "in the next few weeks."

Probably 1% of users will get access to it, with a 20/message a day rate limit. Until early next year.

Re: Learning to Reason with LLMs

#38

oh wow, something you can roughly model as a diy in a base model. so impressive. yawn. at least NVDA should benefit. i guess.

If there's a way to do something like this with Llama I'd love to hear about it (not being sarcastic)

nurture the model have patience and a couple bash scripts

Re: Learning to Reason with LLMs

#39
A lot of skepticism here, but these are astonishing results! People should realize we’re reaching the point where LLMs are surpassing humans in any task limited in scope enough to be a “benchmark”. And as anyone who’s spent time using Claude 3.5 Sonnet / GPT-4o can attest, these things really are useful and smart! (And, if these results hold up, O1 is much, much smarter.) This is a nerve-wracking time to be a knowledge worker for sure.
Post reply on HN