Live data from Hacker News

GPT-5.6

openai.com

741–750 of 1001 posts

Re: GPT-5.6

#741

Where is Gemini in all this? Lately it's not even been in the running. Sir Demis asleep at the wheel? Or Google too scared to release a SOTA model? Or ... maybe Gemini 4 is too good and the NSA is using it to break into systems worldwide ...?

Google pretending they're still far ahead. In reality, far far behind. Google has become "just like another" company, nothing special. Few bright minds, with a lot of overpaid engineers.

Re: GPT-5.6

#742

Earlier quoted context omitted.

There is so much less drama involved with the Codex world. You don't realize how oppressive CC is until you've escaped it. Outages, weird restrictions, degradation, accelerated usage, etc etc etc.

> accelerated usage Can you post more information about this?

The only way to understand those moments where you sometimes Claude acts mysteriously is to be terminally online.

https://www.google.com/search?q=claude+usage+draining+site:w...

Re: GPT-5.6

#743
I find it interesting that no one here has mentioned the increased (usable) context window 258k -> 353k. That's huge, but I wonder if it means we pay long context (2x) for the ones past 272k still.

Re: GPT-5.6

#744
I really appreciate the focus on intelligence WITH token efficiency. I'd like to see that become the trend. Smartest per token metrics. Least tokens to accomplish the task above a certain success level. Most of my tasks would benefit from efficiency / token, but switching models constantly, and trying to guess the right model and effort level takes up too much of my processing.

Re: GPT-5.6

#745
post #474
post #149

Earlier quoted context omitted.

Claude Code fan here... Codex is very good. Sometimes better. The killer feature is price. After 6+ months of exclusive Claude Code usage, I was begrudgingly forced to try Codex once Anthropic rejiggered their limits such that I kept maxing out my $200/mo plan in just a few days. These days I pay both $200/mo plans, and it's just about enough to get me through a week's work (small game studio - infinite code to write…

Genuine question/not a critique-are you actually reviewing all that code or just sending it and hoping for the best? I just can't imagine someone is reading/reviewing that much code every day, but maybe I'm wrong?

Not parent poster but: I probably spend at least 2/3 of my tokens on code review & QA. At least at my workplace, that's the culture.

Re: GPT-5.6

#746

Where is Gemini in all this? Lately it's not even been in the running. Sir Demis asleep at the wheel? Or Google too scared to release a SOTA model? Or ... maybe Gemini 4 is too good and the NSA is using it to break into systems worldwide ...?

The rumor is Gemini 3.5 pro will be released next week.

It had previously been announced (at Google I/O) to be released in June, but it got pushed back. Why? Rumors range from they are having trouble to they scrapped an old framework to pursue a new one and the new model will surpass Fable 5 and be cheaper.

Re: GPT-5.6

#747

Things I have been struggling with Fable over and GPT 5.5, were just solved handily by SOL in a real "thank you, next problem" kind of way. Overall, something that just works is way less wasteful for your usage than struggling back and forth for hours.

[flagged]

Love this comment.

Re: GPT-5.6

#748
post #284

Earlier quoted context omitted.

They reset all usage half an hour ago. It's back to 0% per week and session. No specifically Fable related.

I let it work on two features while I was using mostly GPT 5.6 and it has already consumed 10% of the weekly Fable limit on Max x20. GPT 5.6 on the Pro x5 plan is down to... 100%. It looks like they just reset the usage limits again. And I still have two resets on the bench. Anthropic is going to have to up their game to compete.

What do you mean 'resets on the bench?

Re: GPT-5.6

#749
post #234

GPT-5.6 Sol sets a new SOTA on ARC-AGI-3: 7.8% Sol is the first verified frontier model to ever beat an ARC-AGI-3 game https://arcprize.org/results/openai-gpt-5-6

Seeing the dramatic differences in scores just going from high to xhigh is just another demonstration of the bitter lesson: Just keep scaling search and learning. We are probably going to need a lot more GPUs.

Or a new model. The human brain does far more with far less.

Re: GPT-5.6

#750

Earlier quoted context omitted.

There is so much less drama involved with the Codex world. You don't realize how oppressive CC is until you've escaped it. Outages, weird restrictions, degradation, accelerated usage, etc etc etc.

At least Anthropic doesn't bend down to Pentagon/Trump administration

Perhaps not, but I wouldn't say Anthropic is too far removed.

https://investors.palantir.com/news-details/2025/Anthropic-J...

Post reply on HN