[flagged]
Previewing GPT‑5.6 Sol: a next-generation model
311–320 of 797 posts
Re: Previewing GPT‑5.6 Sol: a next-generation model
#312Re: Previewing GPT‑5.6 Sol: a next-generation model
#313Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
Re: Previewing GPT‑5.6 Sol: a next-generation model
#314Earlier quoted context omitted.
"we can start getting these answers back faster, they end up being more useful." Dude, 10x token speed is going to be absolutely nuts. Half the "parallel subagent workflow" business seems to be driven simply as a means to avoid tapping your thumbs waiting for the infernal robot to finish something. If things come back speedy quick all the time, it should keep up with the "speed of the human" and let me stay focused o…
it also makes the parent brain-dead because all those subtokens are missing from the context thus unable to steer the hyper dimensional context driven generation, and the subagent is dumb as a post so synthesizes something very weedsy while you're specifically attempting to understand the forest
Re: Previewing GPT‑5.6 Sol: a next-generation model
#315Earlier quoted context omitted.
He created Django, what do you mean he's not an engineer? Also 'low-effort??' his posts are extremely in-depth, clearly very thought through with a significant amount of time and energy. Additionally he does perform multifaceted checks across LLMs in many of his other blog posts.
> He created Django, what do you mean he's not an engineer? I specifically said that he is not an ML engineer (emphasis on ML), so I'm not sure what Python web frameworks have to do with anything. > Also 'low-effort??' his posts are extremely in-depth, clearly very thought through with a significant amount of time and energy And yes, low effort. Pelican was low effort, his Fable test was low effort, his HN filter etc…
Re: Previewing GPT‑5.6 Sol: a next-generation model
#316 we expect substantial benefit for legitimate defensive work, while meaningfully constraining prohibited offensive use.
That's literally impossible. Writing an exploit agains a known vulnerability needs the exact same knowledge that defending against the exploit of the same vulnerability.Also just making the model better at code is just making it better to writing offensive code.
Re: Previewing GPT‑5.6 Sol: a next-generation model
#317All of these LLMs are getting better at being at an LLM But GPT-5.5 is as useful an LLM can be; it has solved lemmas I've thought about for a year, it can implement typed STLCs in Rust when I give it a formal grammar, it can help me analyze Postgres planner dumps. It's great at tasks that have short solutions but - they cannot learn based on a project - their long term planning capabilities are worse than worms - the…
> - their internal representations are disgusting compared to JEPA You say this based on a theoretical understanding or did you inspect them?
JEPA gives you interpretability for free.
I have not personally inspected them and my view is maybe a more exaggerated/dramatic claim of those working in the JEPA sphere
Re: Previewing GPT‑5.6 Sol: a next-generation model
#318Re: Previewing GPT‑5.6 Sol: a next-generation model
#319To me that means “it’s an inferior product but marketing dictates we try and hide that.”
And “our most robust safety stack to date. We strengthened protections for higher-risk activity, sensitive cyber requests, and repeated misuse, and spent multiple weeks finding weaknesses, pressure-testing our system, and hardening it against real-world attacks” is of zero value to me at best, and most likely to my detriment (increasing refusals or nerfing utility). Why do providers keep leading with that? Are there customers (besides support ChatGPT chatbot users, maybe??) that ask for this?
Re: Previewing GPT‑5.6 Sol: a next-generation model
#320Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…
For comparison, openrouter says opus 4.8 is ~55 tokens/s and fast mode is ~102. 750 tokens/s for their largest model is going to be nuts