Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

61–70 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#61

    Flagged activity can also trigger account-level review across relevant conversations and risk signals, consistent with our terms and policies around content retention and review. Looking beyond a single conversation helps our systems distinguish persistent malicious behavior from legitimate dual-use security work, where similar technical concepts may appear in very different contexts.
Fascinating!

Every conversation you have with these "more capable" models will be monitored and joined up and then your entire account might one day be tagged as Distiller or Cyber Threat Actor or whatnot. When combined with identity verification (which isn't discussed in this press release), expect people to be falsely flagged and banned from ever using OpenAI models again.

Wish I could find the thread from last week where discussions of exactly this kind of thing were dismissed as daft and outlandish.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#63
post #8

I'm going to pre-register my prediction that GPT-5.6 Sol is significantly behind Claude Fable 5, as evaluated by general consensus once time has passed for people to get familiar with both.

I suspect GPT-5.6 Sol will at-the-least be affordable.

"Affordable" depends on what you need. When a task is able to be achieved by two different calibers of model, it's obviously more cost effective to use the less capable model, in the same way that you wouldn't hire a math PhD to do simple addition.

If what you need is only possible with the more capable model then the "affordability" of the less capable model is sort of irrelevant. If what you need is a novel mathematical proof, it doesn't matter that a high school student is "more affodable". You need the math PhD.

As "old" models get more and more capable, it's going to be an increasingly important skill to be able to adequately recognize when a task requires a frontier model and when it doesn't, so that the less capable (and therefore cheaper) model can be used.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#64
post #5

If it's a new generation why isn't it GPT-6?

It does not introduce incompatibilities with earlier 5.x models? Frontier models are at a point now that there will never be a need for another major version bump, aside from those chasing marketing gimmicks. They are smart enough to adapt.

not true. multimodality is still far from being solved

Re: Previewing GPT‑5.6 Sol: a next-generation model

#65
post #10
post #8

I'm going to pre-register my prediction that GPT-5.6 Sol is significantly behind Claude Fable 5, as evaluated by general consensus once time has passed for people to get familiar with both.

What is this prediction based on?

Fable is allegedly a massive model (estimates between 6-10+ trillion, with a few hundred billion active). If 5.6 is just an incremental upgrade over 5.5 (at the same model size) then it won't be able to fully compete with Fable just yet.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#66
post #5

If it's a new generation why isn't it GPT-6?

It does not introduce incompatibilities with earlier 5.x models? Frontier models are at a point now that there will never be a need for another major version bump, aside from those chasing marketing gimmicks. They are smart enough to adapt.

A major bump will be warranted if/when we can truly separate prompt from data.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#67
> Additionally, we’re introducing a new `ultra` mode that goes beyond the capabilities of a single agent by leveraging subagents to accelerate complex work.

I'm curious about how does this work? Do the subagents also get to use the same tools? Will the client be flooded with tool calls? Why extra pricing for a new "model" when the same thing can happen in the client with more controls?

And if it's an army of subagents, why do they compare it to Fable and Mythos? Those models with similar harness would probably bench better I'm guessing

Re: Previewing GPT‑5.6 Sol: a next-generation model

#68
post #8

I'm going to pre-register my prediction that GPT-5.6 Sol is significantly behind Claude Fable 5, as evaluated by general consensus once time has passed for people to get familiar with both.

Claude will win on "vibes" and it'll be close in coding but considering how incremental Fable is above 5.5 in terms of overall smarts, there's no way 5.6 isn't considerably smarter on the whole.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#69
post #59
post #56

Earlier quoted context omitted.

If you have no need for Anthropic/OpenAI's frontier model capability, you may be better served with an open-weight model that can't be taken away. Edit: > GPT-5 does the job. I bring up DeepSeek V4 Flash a lot on HN, but I want to mention that according to Artificial Analysis, it trades blows with GPT-5 (high) (from August, 2025) [0] [0]: https://artificialanalysis.ai/models/comparisons/deepseek-v4...

Unless you are hosting it yourself on your own infrastructure it absolutely can be taken away.

But you have multiple providers, not just one.
Post reply on HN