Live data from Hacker News

GPT-5.6

openai.com

11–20 of 1001 posts

Re: GPT-5.6

#11
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

It never really mattered (except when codex was very new). If anything, codex's remote session integration is better, so outside of some "ultracode" orchestration bells/whistles where Claude Code is ahead, I think Codex is a better tool.

Re: GPT-5.6

#13
"On Agents’ Last Exam (opens in a new window), an evaluation of long-running professional workflows across 55 fields, GPT‑5.6 Sol sets a new high of 53.6, eclipsing Claude Fable 5 (adaptive reasoning) by 13.1 points. Even at medium reasoning, it beats Fable 5 by 11.4 points at roughly one-quarter the estimated cost. That efficiency extends to smaller models, which are essential to making intelligence more abundant and affordable: GPT‑5.6 Terra and GPT‑5.6 Luna outperform Fable 5 at around one-sixteenth the cost. "

Some pretty big claims and results! Excited to see how it feels during usage.

I use Fable and 5.5 extensively and I still find both have a place in my toolkit, i.e. Fable IS good but it isn't perfect, and it's still better to play them off against each other. I have Fable and 5.5 write plans and have them adversarially review each other's plans.

Having this amount of competition in the coding model space is good for all of us.

Re: GPT-5.6

#16
Most importantly, the cost:

> GPT‑5.6 is priced per 1M tokens across three model sizes: Sol is $5 input / $30 output; Terra is $2.50 input / $15 output; and Luna is $1 input / $6 output.

Just as expensive as Fable 5. But of course, another slot machine upgrade but the costs will keep going up and the open weight models from china will continue to race everyone else to $0.

Looking forward to the next version of GLM, Qwen, Deepseek and Minimax.

Re: GPT-5.6

#17
I haven't tried an OpenAI model for a long time, but with Fable going to API pricing soon this might be enough to get me to try codex.

Re: GPT-5.6

#18
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

I left Claude for Codex months ago. I was an early Claude Code adopter but I have found Codex consistently better since about the February time frame. And far more reliable.

It's more diligent and empirical and results focused, and less creative. It sometimes needs a kick to avoid a Zeno's paradox of incremental steps to get to the goal. But it produces more reliable code with fewer race conditions, unhandled negative cases, etc.

It's also better value from a $$ POV, or at least has been. This fluctuates a bit.

You're also free to use your Codex subscription with other harnesses, like opencode, etc. Unlike Anthropic. Plays better with others.

Re: GPT-5.6

#19
post #9
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

Use a harness that doesn't lock you into a moat, like OpenCode.

Can't use a claude code subscription in another harness though

Re: GPT-5.6

#20
post #9
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

Use a harness that doesn't lock you into a moat, like OpenCode.

FWIW Claude Code works with OpenRouter so you can use any model.
Post reply on HN