Live data from Hacker News

GPT-5.6

openai.com

181–190 of 1001 posts

Re: GPT-5.6

#181
post #142
post #97

Earlier quoted context omitted.

Codex has arguably been better than Claude Code for months now, but it's flown under the radar because it just didn't capture the same viral marketing effect and OpenAI in general has had more optics / PR issues than Anthropic amongst the online developer crowd. I use the word "better" not in the sense that the underlying GPT models are fundamentally smarter or more intelligent, but rather that as a product Codex is…

I’d argue the opposite. I’ve switched back and forth from one to the other and Opus/Fable has been constantly better than any GPT in my daily work. It’s a bit slower but it does the things right, with as little code as possible, some comments where needed. Codex is faster but you always have to correct it because it got something wrong; it writes tons of code ("let me add a small helper") with obvious comments.

I'm not sure how meaningful this is. Fable only just recently become more broadly available, and GPT-5.6 is launching broadly today.

Re: GPT-5.6

#182
post #175

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna, and it just moves further into confusing territory when it comes to naming. The naming convention is especially difficult to decipher depending on what your native language is. Of course a latin language speaker might be able to easily determine oh yeah each one is slightly bigger than the other but I still think it borderlines too confus…

You don’t know what sol means? You don’t understand the difference in sizes between Terra and sol? I’m genuinely asking.

Did you not read the second sentence? Obviously I know what sol is given my first language being Spanish. I'm just speaking in a general sense that it can be confusing for others.

I already know plenty who had no clue what the difference between Terra and Luna would be.

Re: GPT-5.6

#183

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna, and it just moves further into confusing territory when it comes to naming. The naming convention is especially difficult to decipher depending on what your native language is. Of course a latin language speaker might be able to easily determine oh yeah each one is slightly bigger than the other but I still think it borderlines too confus…

it's simple: unless trivial TOIL, always use the highest at ultra max settings.

Re: GPT-5.6

#184
> Instead of requiring developers to script every step or passing every tool response back through the model, Programmatic Tool Calling in the Responses API can filter large amounts of intermediate data, retain only what matters, and adapt its workflow along the way.

this seems very interesting

Re: GPT-5.6

#185
post #2

At this point, they are just changing the decimals to stay relevant and in the news.

Anthropic should be grateful OpenAI did not borrow "Epic" and "Legend".

I expect OpenAI names to be "fabulous", "glorious", "empowered", "delicious" etc.

Re: GPT-5.6

#188
post #87
post #68

The frontier graph on all these benchmark are extremely in favor of 5.6 Sol over Fable, more than the best model comparisons in previous iterations. I'd like to know how cherry-picked this is, and what tests it performed less overwhelmingly in, but I suppose that info is not going to be on this post. If it pans out to be as good as it says, that's great. On the other hand, if this model is not overwhelmingly impressi…

They do disclose that they scored much lower than Fable on SWEBench Pro, which is a pretty high-quality benchmark. I think it's partially just about what they choose to emphasize...

Didn't they also just post about how SWEBench is broken?

Re: GPT-5.6

#189
post #175

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna, and it just moves further into confusing territory when it comes to naming. The naming convention is especially difficult to decipher depending on what your native language is. Of course a latin language speaker might be able to easily determine oh yeah each one is slightly bigger than the other but I still think it borderlines too confus…

You don’t know what sol means? You don’t understand the difference in sizes between Terra and sol? I’m genuinely asking.

Sure—so, is Sol 109.2x better than Terra? Or 1.304x10^6 better?
Post reply on HN