Previewing GPT‑5.6 Sol: a next-generation model
181–190 of 797 posts
Re: Previewing GPT‑5.6 Sol: a next-generation model
#182I think GPT writes code the best. How well will it write in version 5.6? It gives me chills. Recently, I went head-to-head with GPT on nearly 2,000 lines of code, and GPT's solution was superior and faster. I even referenced multiple codebases on GitHub while trying, but they were incomparable to GPT. So using GPT brings both fear and excitement. The fear comes from realizing that this level of code is now the averag…
I am on the opposite camp. Open models are starting to perform better. GPT 5.5 keeps on messing things up. On the contrary, pi + glm + DeepSeek… bliss. Fable was a different kind of beast though. Rip.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#183Here is a trend I'm noticing: - GPT-5 mini costs $0.25/$2 and will be discontinued in December. - GPT-5.4 mini costs $0.75/$4.5 and is supposed to be the replacement. - GPT-5.4 nano costs $0.2/$1.25 and, while it ranks better in benchmarks than GPT-5 mini, it's not even close when you test it in real scenarios. So you're left being forced to go to GPT 5.4 mini if you use 5 mini today. The same thing is happening here…
On Nano "it's not even close when you test it in real scenarios" - what have you seen? What kind of things can GPT-5 Mini handle that GPT-5.4 Nano cannot?
Re: Previewing GPT‑5.6 Sol: a next-generation model
#184Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
[0]: https://openai.com/index/openai-broadcom-jalapeno-inference-...
Re: Previewing GPT‑5.6 Sol: a next-generation model
#185Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
750 tokens/s for their largest model is going to be nuts
Re: Previewing GPT‑5.6 Sol: a next-generation model
#186Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
Re: Previewing GPT‑5.6 Sol: a next-generation model
#187Earlier quoted context omitted.
Yeah, this is the classic silicon valley strategy of selling at a loss and then once they have captured the market inflate prices. See Uber, Netflix, etc.
This is a constantly repeated conspiracy theory and is not true at all. The api costs do increase but aggregate costs per task decrease. The question is: do people need lower intelligence models at all? The answer is a resounding NO! How many people do you see using haiku or sonnet? I see very few and most people default to the latest model and just play with thinking effort. I think three layers are good enough and…
For my use case a model from a year ago is good enough
Re: Previewing GPT‑5.6 Sol: a next-generation model
#188So the next naming scheme might be FTX, Madoff and Enron? :^)
Re: Previewing GPT‑5.6 Sol: a next-generation model
#189When will GPT-5.6 Protomolecule drop? Me and the boys on Eros can't wait to get our hands on it!
Re: Previewing GPT‑5.6 Sol: a next-generation model
#190Earlier quoted context omitted.
For all intents and purposes you'll be able to move an open weight model wherever you want. I really dislike this rhetoric, you sound like the FSF guys who are like "you're not free until you're running coreboot with zero binary blobs". Sure they have a point but also, most people are fine running regular linux.
Unless the US Gov bans inference companies from serving Chinese models to US customers...