Live data from Hacker News

Advancing the price-performance frontier with GPT‑5.6

openai.com

81–90 of 424 posts

Re: Advancing the price-performance frontier with GPT‑5.6

#81
post #30

80% price cut for luna is a very aggressive pricing move makes it by far the best choice for most workloads that do not need bleeding edge intelligence (reminder: luna can be comparable to opus 5!)

Luna is meant to compete with Haiku. What tasks are you seeing it equal Opus on?

Re: Advancing the price-performance frontier with GPT‑5.6

#82

Earlier quoted context omitted.

This is gonna put Sonnet 5 in a really awkward spot.

Luna is comparable to Haiku, not Sonnet.

Totally untrue. Luna and Sonnet 5 are very comparable: https://artificialanalysis.ai/#intelligence

Luna is an extremely strong model.

Re: Advancing the price-performance frontier with GPT‑5.6

#83
post #63

Didn't expect that. Luna pricing is crazy now. I don't think there is anything on the market that competes at this price-performance point. For our production app, OpenAI clearly is the best provider now. Their API is very reliable and has many nice features. The price-performance of the model lineup is incredible. We used open weights model via Fireworks for a long time (e.g. Kimi K2.5). Fireworks is a great provide…

OpenAI's APIs are extremely reliable for sure. I don't even remember when the last incident or downtime was.

Re: Advancing the price-performance frontier with GPT‑5.6

#84
post #9

Making Luna, which was already very cheap and extremely capable, 5x cheaper is crazy. I use Sol at work but Luna at home, and while there's definitely a difference, it doesn't feel like night-and-day. After a year of ever-increasing prices it suddenly feels (between this, Kimi K3, GLM 5.2) that prices are falling again.

> Sol vs Luna > it doesn't feel like night-and-day. I see what you did there. :)

[deleted]

Re: Advancing the price-performance frontier with GPT‑5.6

#86
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

When model intelligence reliably hits 90%-95% of current day knowledge worker tasks, they are going to burn those weight to silicon and we will see another 10X improvement in price/performance frontier.

The dynamic GPU clusters will be used for the 5% of tasks, and pushing out the frontier. Also there will be a set of knowledge tasks that are not done today (because they are too difficult for most knowledge workers), that will start being done in the future.

Re: Advancing the price-performance frontier with GPT‑5.6

#88
post #59

Might just resub. Will experiment with Luna next sessions. 5 h window is not working very well for me. But if I can drop down to Luna at 20-30 % left and comfortably ride out the wave then.. that might just work.

They got rid of the 5hr quota, it's just weekly quotas now.

Re: Advancing the price-performance frontier with GPT‑5.6

#89
post #73
post #50

Earlier quoted context omitted.

Do you have a sense of which tasks benefit from more agents and which don't?

Anything related to reading and interpreting the environment seems to always benefit from the addition of more agents to the search party, assuming you have some rational way to synthesize their results. Taking actions that mutate the environment is a different story. I think this is where you run into diminishing returns very quickly. You generally want one strong agent to act given the results of all the searching…

I definitely think you want the genius model to synthesize everything that rolls up to them.

Re: Advancing the price-performance frontier with GPT‑5.6

#90
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

To be fair we don't really know in terms of prices what's real and what's just investor subsidised attempts at market capture at this point. It could well be OpenAI's attempt to drown Anthropic while they've got the halo product if they feel they've got deeper pockets.

I wouldn't be surprised if they still had some margins since cheaper models are much harder to nail the accurate sizes off, and you still pay 2x for 1M context window.

But if this is even at 400B size it's insanity those inference prices, maybe 10-20% margins, if it's higher I would like to know is it their own chips or maybe they have accurately sized the model to fit on exactly a B300?

Could be a lot of magical things we can only speculate, but from here there likely isn't another 60-70% margin, like I have heard people claim, I would definitely be willing to bet on that.

Could still be a healthy 10-30% margin. Especially with Terra.

Post reply on HN