Live data from Hacker News

GPT-5.6

openai.com

681–690 of 1001 posts

Re: GPT-5.6

#683

Earlier quoted context omitted.

> Avoid generic brevity instructions That part is confusing because it's not like they provide an example of how default GPT-5.6 output compares with GPT-5.5 both with default output and prompted for brevity. Whenever I use such prompts, it's usually because I want the model to give me the gist in a few sentences. I'd be stunned if GPT-5.6 was that concise by default. I would think that could "break" a lot of things…

It sure is suspicious that both Anthropic (adaptive thinking) and OpenAI (Avoid generic brevity instructions) both seem to be suggesting that the best way to improve outcomes is to entirely leave it to them to decide how many tokens get used. I mean, it's true that it would be ideal of this stuff did just get figured out optimally behind the API, but there is definitely an incentive on their side to burn more tokens.

Perhaps the incentive is for variable behavior. When there is low GPU demand, burn more, but reduce when there is contention.

Re: GPT-5.6

#684

We Openly hate OpenAI because they’re not very Open but we secretly hope they win against not-open-at-all Anthropic.

I openly hope the chinese labs distilling them into open weights win.

Tried Xiaomi MiMo v2.5 via opencode today. Since Sonnet 5's release week, Sonnet 4.6 has been feeling like a vegetable, with Sonnet 5 itself being only a little better. MiMo on the other hand feels like Sonnet 4.6 did up until very recently. Absolutely impressive.

In some ways, more impressive than GPT 5.5 with high(!) thinking. GPT says quite some nonsense from time to time; didn't see any sign of this in MiMo so far, which is a pretty wild difference.

Re: GPT-5.6

#685
post #379

Here are 18 pelicans - six each for Luna, Terra and Sol at the six different reasoning effort levels (plus the price to generate each one): https://static.simonwillison.net/static/2026/gpt-5.6-pelican... Or if you want to see some in 3D, OpenAI featured a pelican riding a tricycle, bicycle, pony and another pelican in their livestream this morning: https://www.youtube.com/live/Wq45rvPGNHs?t=1070s

[dead]

Re: GPT-5.6

#686
Just my two cents. I'm on the Plus plan, I ask gpt-5.6 sol / high to analyze a vibe-coded codebase (~50k LoC) and write a plan to make it production ready. It wasn't a great prompt, I just wanted to test it quickly. It ran for ~15min and consumed 95% of my 5h quota (I thought it was gonna crash). The output is excellent but just a heads up that it consumes a lot of quota!

Re: GPT-5.6

#687

We Openly hate OpenAI because they’re not very Open but we secretly hope they win against not-open-at-all Anthropic.

personally I hope any company involved with child slaughter ends up crashing and burning, i say this because both those companies are buddies with the us department of war (who helped annihilate a school the other day)

Re: GPT-5.6

#688

Earlier quoted context omitted.

> Intent understanding This will totally make it brain damaged over a certain tasks. Sort of like the same brain damage that prompted OpenAI project managers to destroy ChatGPT.app today.

> destroy ChatGPT.app today. ... What changed, exactly?

Codex.app is gone and merged into ChatGPT.app. The upgrade process was... messy... Codex's self-update just deleted the Codex.app w/o further instruction. And ChatGPT updater failed the first time while also bricking the prior installed ChatGPT.app.

Seems good/fine once you get through upgrading the app.

Re: GPT-5.6

#689

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna, and it just moves further into confusing territory when it comes to naming. The naming convention is especially difficult to decipher depending on what your native language is. Of course a latin language speaker might be able to easily determine oh yeah each one is slightly bigger than the other but I still think it borderlines too confus…

I love how all the replies to this comment recommend completely different strategies for deciding which model to use.

Re: GPT-5.6

#690

Earlier quoted context omitted.

This is a major reason why I and a number of biologists I've talked to have canceled their anthropic accounts recently. Not working is not working.

It's so absurdly sensitive. It bailed out earlier today working on a TypeScript client for a sensor network API which happens to include some temperature and pH sensors for tanks, which yes, are used for biology experiments. But wow, we're degrees of separation from the actual biology work. It's making it very hard to justify even trying to use Fable. When it works, awesome; it's legitimately good. But I can't trust…

It refused to give me plant care instructions for an ornamental sold at my local Home Depot because it decided it was highly invasive and dangerous to grow in my region.

(It’s not)

Post reply on HN