Live data from Hacker News

GPT-5.6

openai.com

61–70 of 1001 posts

Re: GPT-5.6

#61

CTRL-F: Fable 15 hits Holy shit. They must be feeling very threatened by Fable if they're spending this much energy talking about it in the release notes for their own model.

In the past they received a lot of hate for not comparing to the competition.

Re: GPT-5.6

#62
The meat of the report for SWEs:

SWE-Bench Pro Sol: 64.6% Fable: 80% Opus: 69.2% (!!!!)

So, it still trails Opus, significantly, and is not a next-gen coding model like Mythos/Fable 5.

Disappointing to say the least, but somewhat expected.

Re: GPT-5.6

#65
post #16

Most importantly, the cost: > GPT‑5.6 is priced per 1M tokens across three model sizes: Sol is $5 input / $30 output; Terra is $2.50 input / $15 output; and Luna is $1 input / $6 output. Just as expensive as Fable 5. But of course, another slot machine upgrade but the costs will keep going up and the open weight models from china will continue to race everyone else to $0. Looking forward to the next version of GLM, Q…

That's wrong. GPT 5.6 Sol looks to have the same price as GPT 5.5, apart from a new pricing fee for cache writes. Fable 5 is $10 input / $50 output.

Re: GPT-5.6

#66
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

Literally every top model is identical and anyone saying otherwise is engaging in astrology.

Re: GPT-5.6

#67
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

More literal, less fluid verbally, harder time understanding nuance, more correct code, fewer bugs. Less pretty UI. I switch back and forth but find I have less 'clean up' work with codex; more upfront communication though to properly specify. High hopes for 5.6!

Re: GPT-5.6

#68
The frontier graph on all these benchmark are extremely in favor of 5.6 Sol over Fable, more than the best model comparisons in previous iterations.

I'd like to know how cherry-picked this is, and what tests it performed less overwhelmingly in, but I suppose that info is not going to be on this post.

If it pans out to be as good as it says, that's great. On the other hand, if this model is not overwhelmingly impressive over Fable, I will lose what remaining trust I had in these announcements.

Re: GPT-5.6

#69

Earlier quoted context omitted.

Can't use a claude code subscription in another harness though

You absolutely can; they are not banning anymore. The bigger problem is that subscription versions of the models are way crappier than when the "same" model is hit via API (Bedrock/Vertex) You can also make it not count against extra usage. OpenCode docs show it because Anthropic specifically ambushed them with a PR to remove support so simpletons can't use it easily.

The opencode docs[0] still say otherwise, do you have a source?

[0] https://opencode.ai/docs/providers/#anthropic

Re: GPT-5.6

#70
GPT-5.6 Sol, Terra, and Luna. at this rate GPT-6 will be named after a parking lot and GPT-7 after whatever Elon names his next kid.
Post reply on HN