Live data from Hacker News

GPT-5.6

openai.com

51–60 of 1001 posts

Re: GPT-5.6

#51
post #22

5.6 Terra (mid tier model) as good as Fable on DeepSWE while cheaper than Opus API pricing. Seems like a homerun.

GPT usually performs better on DeepSWE while Claude does better on FrontierCode. These two coding benchmarks are pretty much the only ones right now that's still worth taking a look at imo.

Re: GPT-5.6

#54
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

In my experience, for coding Codex is definitely far ahead of Claude Code, even when using Fable 5 as a model.

Re: GPT-5.6

#55
The cost & output token charts are useful but I wish I could view them more like a 3D surface. Like the CS:APP memory mountain charts.

I wonder how long model size and effort will be a few discrete points instead of continuous.

Re: GPT-5.6

#56
Is any of those comparisons about Pro vs non-Pro (Pro is only available in $100+ plans)? I am curious about that but I think Sol, Terra, Luna are different sizes of it without the Pro part, and I want to know how much worse do I have it on the $20 plan compared to if I upgrade.

Re: GPT-5.6

#57
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

I use both constantly for different things. You don't need to be a one-model Andy

Re: GPT-5.6

#59

CTRL-F: Fable 15 hits Holy shit. They must be feeling very threatened by Fable if they're spending this much energy talking about it in the release notes for their own model.

Apparently it significant outperforms fable on both an intelligence and cost index.

I don’t believe it at all and I don’t think anyone else does either.

Re: GPT-5.6

#60
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

Not sure there's going to be a consensus, but I can tell you that when i have codex review claude-written code, it finds important gaps and fixes. The reverse is also true. Both are powerful, but even better when used in combination
Post reply on HN