Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

181–190 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#182
post #77
post #44

I think GPT writes code the best. How well will it write in version 5.6? It gives me chills. Recently, I went head-to-head with GPT on nearly 2,000 lines of code, and GPT's solution was superior and faster. I even referenced multiple codebases on GitHub while trying, but they were incomparable to GPT. So using GPT brings both fear and excitement. The fear comes from realizing that this level of code is now the averag…

I am on the opposite camp. Open models are starting to perform better. GPT 5.5 keeps on messing things up. On the contrary, pi + glm + DeepSeek… bliss. Fable was a different kind of beast though. Rip.

Yeah, Opus/GPT need multiple rounds of reviews from each other to get to clean auto review. Fable was like, it is done and indeed… crickets in bot comments. ‘No issues’ galore.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#183
post #71

Here is a trend I'm noticing: - GPT-5 mini costs $0.25/$2 and will be discontinued in December. - GPT-5.4 mini costs $0.75/$4.5 and is supposed to be the replacement. - GPT-5.4 nano costs $0.2/$1.25 and, while it ranks better in benchmarks than GPT-5 mini, it's not even close when you test it in real scenarios. So you're left being forced to go to GPT 5.4 mini if you use 5 mini today. The same thing is happening here…

On Nano "it's not even close when you test it in real scenarios" - what have you seen? What kind of things can GPT-5 Mini handle that GPT-5.4 Nano cannot?

We’re using GPT-5-mini in an enterprise data-processing workflow, and we too see that GPT-5.4 nano performs materially worse for our requirements, roughly 30% worse as measured through our test suite.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#184

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

OpenAI also announced two days ago that they're starting to make Cerebras style chips themselves [0], will be interesting to see how fast SotA model inference will be by the end of the year.

[0]: https://openai.com/index/openai-broadcom-jalapeno-inference-...

Re: Previewing GPT‑5.6 Sol: a next-generation model

#185

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

For comparison, openrouter says opus 4.8 is ~55 tokens/s and fast mode is ~102.

750 tokens/s for their largest model is going to be nuts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#186

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

Yep this is a glimpse into the future of 500+ t/s, which is in my opinion the next big thing that validates Jevon's paradox (the models are already smart enough)

Re: Previewing GPT‑5.6 Sol: a next-generation model

#187

Earlier quoted context omitted.

Yeah, this is the classic silicon valley strategy of selling at a loss and then once they have captured the market inflate prices. See Uber, Netflix, etc.

This is a constantly repeated conspiracy theory and is not true at all. The api costs do increase but aggregate costs per task decrease. The question is: do people need lower intelligence models at all? The answer is a resounding NO! How many people do you see using haiku or sonnet? I see very few and most people default to the latest model and just play with thinking effort. I think three layers are good enough and…

Do I need the most intelligent model to generate boilerplate code, which is my main usage for AI? Resounding No.

For my use case a model from a year ago is good enough

Re: Previewing GPT‑5.6 Sol: a next-generation model

#189
post #11

When will GPT-5.6 Protomolecule drop? Me and the boys on Eros can't wait to get our hands on it!

Musk steals Dario and they both train Epic on Mars. US Space Force promptly finds oil on Mars and launches an armada in the next window. In the meantime rocks painted black drop on Mar-a-Lago.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#190
post #180

Earlier quoted context omitted.

For all intents and purposes you'll be able to move an open weight model wherever you want. I really dislike this rhetoric, you sound like the FSF guys who are like "you're not free until you're running coreboot with zero binary blobs". Sure they have a point but also, most people are fine running regular linux.

Unless the US Gov bans inference companies from serving Chinese models to US customers...

good luck doing it to inference companies in singapore or the netherlands. or one of the decentralized networks that dont look useful right now. the world is already sick of america acting like it can do whatever and force their rules on the rest of us.
Post reply on HN