Earlier quoted context omitted.
But it’s irrelevant. 750 tokens/s on a full frontier model is useful. 15000 poor quality tokens is much less useful no matter how much scaffolding you put around it.
I’ve been using 1,000 t/s on a near frontier model for a month now. It’s very useful for agentic coding. It does require new approaches for me personally since I get a lot less time to think or read its output.
Previewing GPT‑5.6 Sol: a next-generation model
761–770 of 797 posts
Re: Previewing GPT‑5.6 Sol: a next-generation model
#762Earlier quoted context omitted.
“Smart enough” really depends on how many other people have encountered a problem close enough to yours and solved it somewhere on the open internet, IMO. Most of the frontier models can, when prompted and tooled correctly, do a lot of “reasoning” tasks that amount to resolving how the user has explained a particular widely known paradigm. The more difficult and obscure the issues you provide them with, the faster yo…
This may have been the case one year ago, but with contemporary models such as Opus, I run into this less often.
Fable was actually a lot better at not doing it in the couple of brief windows I got to use it.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#763Earlier quoted context omitted.
Maybe true, but if you're using an LLM to do some real world work, do you want it to have some abstract notion of intelligence, or do you want it to actually do the job you assigned it?
I want it to not murder or opress lots of people by mistake
Re: Previewing GPT‑5.6 Sol: a next-generation model
#764Re: Previewing GPT‑5.6 Sol: a next-generation model
#765Earlier quoted context omitted.
Yeah, we'll share a lot more details and evals when we can release GPT-5.6 widely. We focused on cyber (and bio) here to help explain why it's being held back for now. We would have loved to launch it to everyone - it's the best coding model I've ever used - and we plan to do so as soon as we can ('coming weeks'). (I work at OpenAI.)
So now have to be worried that I'm going to killed by an AI designed nerve agent that someone has cooked up in their shed? FFS. I hate this world so much. I wish I could just flip a switch and never have to hear about or have anything to do with AI ever again. Do you ever stop to think about the horrific dystopia you and your acolytes are creating?
Re: Previewing GPT‑5.6 Sol: a next-generation model
#766How are they able to compare with Fable when Fable was only available for three days?
Terminalbench numbers are publicly available. What is more interesting, why is that the only benchmark they highlight. Maybe 5.6 isn’t that far ahead of Fable 5 in DeepSWE and FrontierCode (which I consider the most useful and close to my evals + subjective experience)…
Re: Previewing GPT‑5.6 Sol: a next-generation model
#767I saw they are placing this model above Mythos and Fable. Interesting to see how good it's going to compare. I'd really like to see other companies like Chinese ones compete at this level. Pricing on GPT 5.5 is already super high and having more competition can only help :)
Re: Previewing GPT‑5.6 Sol: a next-generation model
#768Earlier quoted context omitted.
But you'd still need code if you need something done in a consistent way.
Not necessarily. Consider a human assistant who performs repetitive tasks at an acceptable cost and accuracy while dealing with edge cases often autonomously.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#769Earlier quoted context omitted.
That's literally the official FSF position. https://www.fsf.org/resources/hw > For example: the Free Software Foundation only purchases desktop machines which support Libreboot, and Thinkpad X200 and X60 laptops with Libreboot. All desktops and servers we buy are KGPE-D16 motherboards, which are supported by Libreboot. As a result, all of the workstations used by the FSF staff have a free BIOS. https://www.gnu.org/di…
> They are also the reason you can buy a computer meeting those requirements The latest libreboot-compatible laptop I could find, at https://libreboot.org/docs/install/t480.html , is from 2018 -- not sure if that would still be available?
https://tehnoetic.com/laptops/tet-t400s
Libreboot actually isn't free anymore by FSF standards because it has binary blobs for the Intel Management Engine. Neither is Coreboot which has blobs for many other things. Most modern computers cannot boot without binary blobs for ME, which is why GNU Boot was created.
If you want a modern/new computer, Purism sells them with the management engine "neutralized" but it's still not free by FSF standards because they haven't bypassed it entirely.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#770Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
For comparison, openrouter says opus 4.8 is ~55 tokens/s and fast mode is ~102. 750 tokens/s for their largest model is going to be nuts
Emphasis on "up to". Imagine whatever limited situation (e.g. pre-cached query) and that will probably be the only time it hits 750 tokens/sec.