Live data from Hacker News

Advancing the price-performance frontier with GPT‑5.6

openai.com

51–60 of 424 posts

Re: Advancing the price-performance frontier with GPT‑5.6

#51
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

This is gonna put Sonnet 5 in a really awkward spot.

Sonnet and Haiku were already in an awkward spot, likely by design.

Anthropic's big marketing push this year has been entirely focused on getting people to use Opus via a Claude Code subscription, to the point that Sonnet is almost viewed as the poor man's alternative, and from what I've seen, almost nobody uses it.

Actually, here's an interesting project for all the vibe coders looking for their next front page post: scrape a ton of commits from GitHub with Co-Authored-By: Claude and figure out what the percentage split between Opus/Fable/Sonnet is. I'm willing to bet it's less than 10% Sonnet.

Re: Advancing the price-performance frontier with GPT‑5.6

#54
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

it's 80% less cost, not 80% in efficiency gains, could be that Luna was overpriced to begin with, we don't have much info on the models themselves.

Assuming the efficiency gains are real, I feel like something has to give, maybe worse quality due to aggressive quantization/kv cache compression?

Re: Advancing the price-performance frontier with GPT‑5.6

#55

"Half the money I spend on advertising is wasted; the trouble is I don't know which half." -John Wanamaker This applies even more strongly to model choosing. I know for a fact that majority of my work doesn't require a very strong model, but separating the trivial and non-trivial tasks is a famously hard problem (if at all decidable).

Cosmically apt username given the substance of this comment.

Re: Advancing the price-performance frontier with GPT‑5.6

#56
post #42

Looks like I might have a reason to use something other than Deepseek V4 Flash.

I was curious so I photoshopped DSV4 Flash into the graph:

https://files.catbox.moe/csxl32.png

(2 cents to run AA index, score 40)

Looks like OpenAI broke the pareto frontier on the trust-me-bro benchmarks!

(One has to wonder if they used any of the neat tricks from the DSV4 paper :)

Re: Advancing the price-performance frontier with GPT‑5.6

#57
post #15
post #7

Earlier quoted context omitted.

imagine writing that on your resume > reduced inference cost by 20 percent saving company x billion dollars per month

In this case, and I don't mean this critically, I guess it would technically be, "Instructed model to find efficiencies... reducing inference cost by 20% saving company x billion dollars per month." I have no doubt that further work was required to enable this, but it's still very cool to be possible to say that.

[deleted]

Re: Advancing the price-performance frontier with GPT‑5.6

#58
post #33
post #7

Earlier quoted context omitted.

imagine writing that on your resume > reduced inference cost by 20 percent saving company x billion dollars per month

Where are you going to apply to with that resume that’s a step up from your current job though?

lots of places, actually. not everyone wants to be attached to the Silicon Valley culture, and that line alone will guarantee practically any workplace. that person is going to find out what work-life balance is :)
Post reply on HN