Earlier quoted context omitted.
Yep this is a glimpse into the future of 500+ t/s, which is in my opinion the next big thing that validates Jevon's paradox (the models are already smart enough)
I think the glimpse that is there will be exclusive access. So much for the open in openAI. If this technology really transforms society in the ways expected with inequality an unavoidable consequence equal access should be required like internet access was (isp can’t give preference to specific user traffic)
Previewing GPT‑5.6 Sol: a next-generation model
271–280 of 797 posts
Re: Previewing GPT‑5.6 Sol: a next-generation model
#272Earlier quoted context omitted.
For comparison, openrouter says opus 4.8 is ~55 tokens/s and fast mode is ~102. 750 tokens/s for their largest model is going to be nuts
Using gpt-5.4-mini in off-peak hours already feels like super-speed to me. That's probably no more than 100-150 tk/s. I can't imagine 750! I've always eyed Cerebras but never had a use for it that would justify paying for the API directly. Although now that I think about it, trying out the API would probably cost less than a subscription for a month...
If you have a subscription it's a different pool of usage.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#273Re: Previewing GPT‑5.6 Sol: a next-generation model
#274Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
"we can start getting these answers back faster, they end up being more useful." Dude, 10x token speed is going to be absolutely nuts. Half the "parallel subagent workflow" business seems to be driven simply as a means to avoid tapping your thumbs waiting for the infernal robot to finish something. If things come back speedy quick all the time, it should keep up with the "speed of the human" and let me stay focused o…
Re: Previewing GPT‑5.6 Sol: a next-generation model
#275"Next generation model" If it was the next generation, why isn't it a major version change..?
AFAIK there is no difference between "generation" and "version". Version naming/numbering depends on how good it turns out to be, and competition. If the competition releases something then you need to push something out too. Calling it 5.6 creates the least possible expectations, and therefore more potential for positive feedback. The Sol/Terra/Luna naming is interesting. I wonder what Anthropic are considering for…
Re: Previewing GPT‑5.6 Sol: a next-generation model
#276Did GPT-5.6 Sol Ultra decide the terrible colors for the benchmark graphs?
OpenAI's plot design has been consistently awful and inaccessible, it seems like they're optimizing for something other than readability because I find it hard to believe they aren't putting in any effort for such major announcements. If the colors have to be awful they should at least differentiate with marker shapes or line dashes.
At least it isn't as bad as the stacked bar chart where the 50-something bar was higher than the 60-something bar.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#277Earlier quoted context omitted.
Unless you are hosting it yourself on your own infrastructure it absolutely can be taken away.
For all intents and purposes you'll be able to move an open weight model wherever you want. I really dislike this rhetoric, you sound like the FSF guys who are like "you're not free until you're running coreboot with zero binary blobs". Sure they have a point but also, most people are fine running regular linux.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#278I do not like the fact that this forces people to remember one more hierarchy of "Sol vs Terra vs Luna". OpenAI was supposed to simplify their naming since at least 2025.
The Sun is bigger than the Earth which is bigger than the Moon.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#279Sol and 5.5 pro are in parity at $5 input / $30 output. What I'm inferring from this is that: - model weight size didn't change, and this is mostly a result of better model architecture and scaled up RL - better hardware utilization and and they're making better margins OR - worse hardware utilization and they're okay with digging into their margins.
I think you meant 5.5.
I agree it is probably the same size model. It's probably exactly built on top of 5.5, just with more training, or else they would have bumped the version number to 6.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#280I didn't know that I was color blind, but thanks to those charts, I think I need to see a doctor... I mean, you can read them even without the colors, but who on earth thought that those are a good set of colors? Oh, I forgot it was probably someone on 'Sol'.
I'm not colorblind and I was depending on the textual context implying Sol was better than Terra. I had to zoom in quite far to actually differentiate between the colors.
If they insist on terrible colors would it be so hard to differentiate by marker shape or line dashing too?