Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

311–320 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#312
post #282
post #259

Earlier quoted context omitted.

Given the expectations everyone has created GPT-6 has to pretty much be AGI.

What is your definition of AGI that the current LLMs don't fit?

Autonomously Generating Income (which is why it will never be released to the general public)

Re: Previewing GPT‑5.6 Sol: a next-generation model

#313

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

Does the Cerebras variant offer input caching and corresponding discounts? Last I checked Cerebras would not cache or would cache but not give discounts for the cached input, making it impractical for agentic use and multiturn conversations.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#314

Earlier quoted context omitted.

"we can start getting these answers back faster, they end up being more useful." Dude, 10x token speed is going to be absolutely nuts. Half the "parallel subagent workflow" business seems to be driven simply as a means to avoid tapping your thumbs waiting for the infernal robot to finish something. If things come back speedy quick all the time, it should keep up with the "speed of the human" and let me stay focused o…

it also makes the parent brain-dead because all those subtokens are missing from the context thus unable to steer the hyper dimensional context driven generation, and the subagent is dumb as a post so synthesizes something very weedsy while you're specifically attempting to understand the forest

You have an agent spawn the agents for you! You can ask Claude to do it for you, he is happy to use sonnet when you ask for grok and opus high when you ask for deepseek.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#315
post #255

Earlier quoted context omitted.

He created Django, what do you mean he's not an engineer? Also 'low-effort??' his posts are extremely in-depth, clearly very thought through with a significant amount of time and energy. Additionally he does perform multifaceted checks across LLMs in many of his other blog posts.

> He created Django, what do you mean he's not an engineer? I specifically said that he is not an ML engineer (emphasis on ML), so I'm not sure what Python web frameworks have to do with anything. > Also 'low-effort??' his posts are extremely in-depth, clearly very thought through with a significant amount of time and energy And yes, low effort. Pelican was low effort, his Fable test was low effort, his HN filter etc…

We at HN: https://xkcd.com/2501/ to basically say that I think you might be considering low-effort what’s actually an attempt at simplifying - which is arguably higher effort

Re: Previewing GPT‑5.6 Sol: a next-generation model

#316

    we expect substantial benefit for legitimate defensive work, while meaningfully constraining prohibited offensive use.
That's literally impossible. Writing an exploit agains a known vulnerability needs the exact same knowledge that defending against the exploit of the same vulnerability.

Also just making the model better at code is just making it better to writing offensive code.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#317
post #296

All of these LLMs are getting better at being at an LLM But GPT-5.5 is as useful an LLM can be; it has solved lemmas I've thought about for a year, it can implement typed STLCs in Rust when I give it a formal grammar, it can help me analyze Postgres planner dumps. It's great at tasks that have short solutions but - they cannot learn based on a project - their long term planning capabilities are worse than worms - the…

> - their internal representations are disgusting compared to JEPA You say this based on a theoretical understanding or did you inspect them?

Look at VLM mechanistic interpretability papers vs just pca on JEPA trained weights.

JEPA gives you interpretability for free.

I have not personally inspected them and my view is maybe a more exaggerated/dramatic claim of those working in the JEPA sphere

Re: Previewing GPT‑5.6 Sol: a next-generation model

#319
“ Terra has competitive performance to GPT‑5.5 [while being 2x cheaper]…”

To me that means “it’s an inferior product but marketing dictates we try and hide that.”

And “our most robust safety stack to date. We strengthened protections for higher-risk activity, sensitive cyber requests, and repeated misuse, and spent multiple weeks finding weaknesses, pressure-testing our system, and hardening it against real-world attacks” is of zero value to me at best, and most likely to my detriment (increasing refusals or nerfing utility). Why do providers keep leading with that? Are there customers (besides support ChatGPT chatbot users, maybe??) that ask for this?

Re: Previewing GPT‑5.6 Sol: a next-generation model

#320

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

For comparison, openrouter says opus 4.8 is ~55 tokens/s and fast mode is ~102. 750 tokens/s for their largest model is going to be nuts

What about 15k tokens per second? [0] I remember looking at this earlier in the year and it being so fast that it feels fake. And, yes, this model is old - but still awesome for what it is.

[0] https://chatjimmy.ai/

Post reply on HN