Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

281–290 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#281

Earlier quoted context omitted.

> how many other people have encountered a problem close enough to yours and solved it somewhere on the open internet I'm 100% sure that all our web, cc, codex or whatsoever sessions are used in the training, RL or either both. This makes the size of the universe models know about at least one order of magnitude bigger than the open internet.

I get how this is a trueism now but I never really understood why it would be useful to scrape cc/codex sessions for training. The relative amount of human input for that is so low (isn't that why they are so loved and used?), how could it actually be useful to them? Wouldn't you wanna focus on people not using it?

Because you provide them with the "problem" and the "solution" and once you have both you can scale your RL pipeline.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#283

Earlier quoted context omitted.

> how many other people have encountered a problem close enough to yours and solved it somewhere on the open internet I'm 100% sure that all our web, cc, codex or whatsoever sessions are used in the training, RL or either both. This makes the size of the universe models know about at least one order of magnitude bigger than the open internet.

I get how this is a trueism now but I never really understood why it would be useful to scrape cc/codex sessions for training. The relative amount of human input for that is so low (isn't that why they are so loved and used?), how could it actually be useful to them? Wouldn't you wanna focus on people not using it?

It's more useful as a set of feedback on the model results. You can do sentiment analysis on the user responses to see if they found the model results useful/frustrating/etc and use that to guide future training

Re: Previewing GPT‑5.6 Sol: a next-generation model

#284
post #37

Earlier quoted context omitted.

It’s the same as the SaaS model. Price keeps going up, and to justify it they keep forcing you to upgrade to new versions with features that nobody asked for.

“More intelligence” is the new feature. Almost everyone is asking for this. Citation: have you looked at OAI and Anthropic’s customer growth numbers?

Every use case of every customer doesn’t need more intelligence. I’m willing to bet that the vast majority will be perfectly fine running on “low intelligence” at a cheap price forever.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#285
post #8

I'm going to pre-register my prediction that GPT-5.6 Sol is significantly behind Claude Fable 5, as evaluated by general consensus once time has passed for people to get familiar with both.

I’m countering this prediction by stating that Fable and Sol will be somewhat similar - this has always been the trend and I see no reason why this should stop now.

OpenAI may have a model in the works that is similar next-gen size and architecture to Fable, but this isn't necessarily it. I'd guess that 5.6 was more of a hasty reaction to Mythos - same base model (same size, same price) as 5.5 but with additional post-training to make it more competitive with Mythos/Fable in some benchmarks.

Mythos/Fable is supposedly next generation in size vs Opus, and is rumored to have some architectural innovation in terms of dynamic routing/compute, possibly only fully enabled with Fable which at $10/50 is still twice the price of Sol 5.6's $5/30, but a big reduction from Mythos preview which had been an astronomical $30/150 possibly due to the dynamic routing not yet having been enabled.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#286

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

> I can think of the tedious task of finding certain functionality within a codebase. I usually can't beat an AI agent harness at this task today.

Yup, I remember "racing" the AIs to figure things out in codebases just a year ago. Today, I have no chance. Whether it is due to degraded reasoning capabilities on my part or better models, I don't know.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#287
post #249
post #4

"Next generation model" If it was the next generation, why isn't it a major version change..?

If they called it 6.0 and it wasn't AGI, you'd see a lot of complaining here too

What is AGI? (I know what the shortcut expands to, I'm curious about your definition. Don't the current models fit?)

Re: Previewing GPT‑5.6 Sol: a next-generation model

#288

People where mocking EU for regulations and now this is happening in the US. I know that Europe is behind in AI but still...

Are cyberweapons/cyberattacks "munitions"? if so, then isn't a machine capable of producing those munitions also itself a munition? I don't think you can put this down to "orange man bad" or "regulations", we're dealing with a genuinely groundbreaking technology with clear military applications

Re: Previewing GPT‑5.6 Sol: a next-generation model

#289

Here is a trend I'm noticing: - GPT-5 mini costs $0.25/$2 and will be discontinued in December. - GPT-5.4 mini costs $0.75/$4.5 and is supposed to be the replacement. - GPT-5.4 nano costs $0.2/$1.25 and, while it ranks better in benchmarks than GPT-5 mini, it's not even close when you test it in real scenarios. So you're left being forced to go to GPT 5.4 mini if you use 5 mini today. The same thing is happening here…

> stay with the models we actually want

If you want control over the models you use, you have to self-host.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#290
Is this a new pre training run independent of 5.5s or post trained on it with Cerebras support and a rebrand of Pro mode at more usable speeds as Sol? The latter seems more likely to me, especially as 5.5 scales very well across its modes so separate branding could make sense, but I don’t see any clear information either way.
Post reply on HN