Live data from Hacker News

GPT-5.6

openai.com

191–200 of 1001 posts

Re: GPT-5.6

#191

CTRL-F: Fable 15 hits Holy shit. They must be feeling very threatened by Fable if they're spending this much energy talking about it in the release notes for their own model.

Downvoted comment but I did find this comparison aggressive and tacky.

Re: GPT-5.6

#192

Earlier quoted context omitted.

holy moly it's in THREE dimensions! AGI solved

So it's failing epically because it generated a tricycle instead of a bicycle?

The livestream presenter goofed and said the test is typically a tricycle, so I wonder if that's just a coverup.

Re: GPT-5.6

#193

CTRL-F: Fable 15 hits Holy shit. They must be feeling very threatened by Fable if they're spending this much energy talking about it in the release notes for their own model.

Anthropic fumble of Fable's release will go down in the history books, makes sense for OpenAI to run with it.

Re: GPT-5.6

#194

Will this run on Cerebas? I'm really looking forward to that.

Sam Altman confirmed during the initial limited release that Sol will run on Cerebras at 750 tok/sec.

Re: GPT-5.6

#195

"GPT‑5.6 is available starting today across ChatGPT, Codex, and the OpenAI API. The rollout is starting globally now and will continue gradually toward full availability over the next 24 hours."

I am on Plus subscription and see Terra and Luna in Codex, but no sign of Sol. Will it be available only on Pro plans?

Re: GPT-5.6

#196
post #175

Earlier quoted context omitted.

You don’t know what sol means? You don’t understand the difference in sizes between Terra and sol? I’m genuinely asking.

Did you not read the second sentence? Obviously I know what sol is given my first language being Spanish. I'm just speaking in a general sense that it can be confusing for others. I already know plenty who had no clue what the difference between Terra and Luna would be.

My first instinct was Sol > Luna > Terra, since Sol is the farthest away, then Luna, and Terra is the closest. Size was not my first instinct. Or should Terra be the best model because its closest to people, then Luna because there have been people on it, then Sol be the worst because no human has been there?

Re: GPT-5.6

#197

Earlier quoted context omitted.

It seems like the way brevity instructions have changed is mis-aligned with how most people would expect to use them or are currently using them. Here's the example they give: > Instead of asking for the shortest possible answer, replace brevity instructions with prioritization: > Lead with the conclusion. Include the evidence needed to support it, any material caveat, and the next action. Omit secondary detail and r…

Replace 2 word instruction ('be concise') with a 38 word instruction. Human can no longer be concise when asking for a few sentences instead of 20 paragraphs of BS they don't want to read when all they want is a summary to verify the general direction of the prompt-work before digging into the details. such progress!

I don't know how intentional it is / was, but LLMs in general just love to hear themselves talk!

Re: GPT-5.6

#198
post #84

Earlier quoted context omitted.

SWE-Bench pro is pretty much useless now even though many ppl still look at it. OpenAI published a report yesterday saying so as well. Only look at DeepSWE and FrontierCode right now for coding imo.

Amazing, a company that does poorly in a benchmark says that benchmark is useless...

SWE-bench series just aren't that great by today's standard, even Anthropic previously stated Claude had memorized solutions for the non Pro version of the benchmark, I suspect the recent increase in the score for the Pro version probably also had similar behaviors.

But anyway, I think it's pretty useless to look at SWE Bench's now when other way better benchmarks exist.

Re: GPT-5.6

#199
I'm disappointed these models continue to be closed source and so expensive.

Open weight models being 10x or more cheaper is just so much more of an unlock than incremental gains for me.

Re: GPT-5.6

#200
post #175

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna, and it just moves further into confusing territory when it comes to naming. The naming convention is especially difficult to decipher depending on what your native language is. Of course a latin language speaker might be able to easily determine oh yeah each one is slightly bigger than the other but I still think it borderlines too confus…

You don’t know what sol means? You don’t understand the difference in sizes between Terra and sol? I’m genuinely asking.

That isn't what "genuinely asking" looks like, you're criticizing using "questions" as cover. It isn't subtle, nor is it constructive.

I agree with them, Sol, Terra, and Luna are confusing names. They mean the same thing as GPT-5.6-Max, GPT-5.6-Plus, and GPT-5.6-Fast but require base knowledge for an analogy.

It feels like it was adding by the marketing department.

Post reply on HN