CTRL-F: Fable 15 hits Holy shit. They must be feeling very threatened by Fable if they're spending this much energy talking about it in the release notes for their own model.
GPT-5.6
61–70 of 1001 posts
Re: GPT-5.6
#62SWE-Bench Pro Sol: 64.6% Fable: 80% Opus: 69.2% (!!!!)
So, it still trails Opus, significantly, and is not a next-gen coding model like Mythos/Fable 5.
Disappointing to say the least, but somewhat expected.
Re: GPT-5.6
#63Re: GPT-5.6
#64Wow, the "Agents' Last Exam" graph looks unreal!
Re: GPT-5.6
#65Most importantly, the cost: > GPT‑5.6 is priced per 1M tokens across three model sizes: Sol is $5 input / $30 output; Terra is $2.50 input / $15 output; and Luna is $1 input / $6 output. Just as expensive as Fable 5. But of course, another slot machine upgrade but the costs will keep going up and the open weight models from china will continue to race everyone else to $0. Looking forward to the next version of GLM, Q…
Re: GPT-5.6
#66Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?
Re: GPT-5.6
#67Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?
Re: GPT-5.6
#68I'd like to know how cherry-picked this is, and what tests it performed less overwhelmingly in, but I suppose that info is not going to be on this post.
If it pans out to be as good as it says, that's great. On the other hand, if this model is not overwhelmingly impressive over Fable, I will lose what remaining trust I had in these announcements.
Re: GPT-5.6
#69Earlier quoted context omitted.
Can't use a claude code subscription in another harness though
You absolutely can; they are not banning anymore. The bigger problem is that subscription versions of the models are way crappier than when the "same" model is hit via API (Bedrock/Vertex) You can also make it not count against extra usage. OpenCode docs show it because Anthropic specifically ambushed them with a PR to remove support so simpletons can't use it easily.