Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

531–540 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#531
post #137

Earlier quoted context omitted.

There's a lot of tacit knowledge in programming. -Why do you cut API boundaries this way? -Why do you change the order of struct fields? -Why do you deliberately insert padding? Most of it depends on the background and context. Sometimes you add it, sometimes you don't. To understand this tacit knowledge, you need access to senior developers. But their attitude often depends on how promising the student is and what b…

What do you think of modern open-source codebases presently available to the public? Is closed-source/proprietary code that much better?

Open source is much better. Closed source is mostly considered 'done' as long as it just works.

One is a 'craft,' the other is 'survival for delivery.'

Re: Previewing GPT‑5.6 Sol: a next-generation model

#532
post #379

“ Terra has competitive performance to GPT‑5.5 [while being 2x cheaper]…” To me that means “it’s an inferior product but marketing dictates we try and hide that.” And “our most robust safety stack to date. We strengthened protections for higher-risk activity, sensitive cyber requests, and repeated misuse, and spent multiple weeks finding weaknesses, pressure-testing our system, and hardening it against real-world att…

>> Terra has competitive performance to GPT‑5.5 [while being 2x cheaper]… > To me that means “it’s an inferior product but marketing dictates we try and hide that.” I interpret this to mean you're about to get today's mainline performance at a fraction of the price.

If that was the case they would have said "equals" or "matches". Instead they say "competitive", as in win-some lose-some.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#534

Earlier quoted context omitted.

I find the Codex usage super generous (but on the $200 plan, I also have the Claude $200 plan). I can run xhigh with subagents pretty much all my waking hours if I want to. If I turn on speed (1.5x) I will hit the 5 hour limit sometimes. I prefer Claude's vibe over 5.5 but 5.5 seems much less lazy. I'm sure it depends a lot on tasks and prompt strategy though.

I easily burn through 3 $200 plans in less than a week. I am often using 4-6 sessions at once and do run overnight goals though typically 2 at once. Almost never use fast. Claude plans are more generous now by about 2-3x but Anthropic slowed their tps a month or so ago so you’re not getting the speed. It’s flip flopped, Codex tightened it significantly recently and used to be more generous. I do split between work, p…

This has been my experience as well, at least for the last few weeks. Codex 5.5 is the better planner and coder across big projects, but Opus is fine, though my Claude 5 hour window lasts ~2x longer than Codex. So I’ll sometimes use an orchestrator/worker skill to spread the load.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#535

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

I saw videos of coding with Mimo-V2.5-Pro UltraSpeed, which is advertised at 1,000 tokens/s, which is very impressive.:

https://www.bilibili.com/video/BV1fME16uEW7

If the time-to-first-token latency also greatly improved, this could be very useful for end-to-end in controls, like autonomous driving for example.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#536

Earlier quoted context omitted.

At a certain rate we will be able to move towards continuous / real-time inference systems. The discrete, turn based solutions are quite confining with how they must be trained. Continuous and real-time would fundamentally alter the domain. From an information theory perspective we are still in dial-up territory with regard to the actual information rate. 750 tokens per second would be a really bad dialup connection.…

And more importantly those 10 million tokens/s should cost fractions of a penny. Tokens need to be dirt cheap so I hope they build out massive solar+battery powered data centers asap.

No anything but wasteful, weak, expensive, environmentally harmful solar. Nuclear is the only path forward for superior energy production, at least until we figure out fusion.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#537

Earlier quoted context omitted.

So we just bend down then?

Unless you work at OpenAI/Anthropic/etc., you are not a part of the "we". If you're asking what the average person can do, then the civic perogative is political action to help elect more AI-cognizant leaders.

There are none. People forget the other side wanted to shut down research entirely, not just release. No idea why people think the other side would have been any better, it would have been even worse. On top of that, anthropic got exactly what they wanted.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#539
post #77

Earlier quoted context omitted.

I am on the opposite camp. Open models are starting to perform better. GPT 5.5 keeps on messing things up. On the contrary, pi + glm + DeepSeek… bliss. Fable was a different kind of beast though. Rip.

Every time I use opus these days I go shut up... you are not fable.. Hard to imagine how just three days with it changed how I saw LLM use.

I'm reluctantly starting to feel grateful that I went camping right over the window that Fable was out.
Post reply on HN