Live data from Hacker News

Advancing the price-performance frontier with GPT‑5.6

openai.com

121–130 of 424 posts

Re: Advancing the price-performance frontier with GPT‑5.6

#121
post #35

This feels like the dialup->broadband transition to me. I was already a huge proponent of Luna for things like deep research. Being able to run 5x more for the same cost is simply bananas. We are already running 10 parallel agents for hypothesis generation. I cannot imagine 50. The statistics become much more interesting & powerful when you can run so many samples of the exact same prompt+model without breaking the b…

Very interesting. Can you share more about your hypothesis/research pipeline? I have been using Sol for those types of task because I figured you'd need more reasoning for getting good ideas, but maybe quantity > quality at a certain point?

Re: Advancing the price-performance frontier with GPT‑5.6

#122
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

They have no (other) equivalent to nano, so it makes sense that it’s much cheaper now. It may have been better, but it was also hell of a lot more expensive.

Re: Advancing the price-performance frontier with GPT‑5.6

#123

"Half the money I spend on advertising is wasted; the trouble is I don't know which half." -John Wanamaker This applies even more strongly to model choosing. I know for a fact that majority of my work doesn't require a very strong model, but separating the trivial and non-trivial tasks is a famously hard problem (if at all decidable).

Exactly. I haven't reached the "let 1000 agents bloom" mode yet, so currently I'm spending real headspace managing agents doing work, and that work is all important, so why "settle" for sub-frontier models for that work? Maybe I'll get there for non-coding work.

Re: Advancing the price-performance frontier with GPT‑5.6

#124
post #47

I generally just check the Price/Performance graph on Openrouter: https://openrouter.ai/rankings#performance#benchmarks . Activate the "Show Pareto" toggle on the right. I was still using GLM-5.2 in my personal projects, but this just made Luna a very easy choice.

The official doc says, Luna = Previous Nano models, kind of. Is it really good at coding?

Depends on the reasoning effort, see https://deepswe.datacurve.ai (add Luna via model selection drop down, if it's not shown by default)

Re: Advancing the price-performance frontier with GPT‑5.6

#125
post #30

80% price cut for luna is a very aggressive pricing move makes it by far the best choice for most workloads that do not need bleeding edge intelligence (reminder: luna can be comparable to opus 5!)

Luna is meant to compete with Haiku. What tasks are you seeing it equal Opus on?

Per Artificial Analysis:

- Haiku: 30 points

- Luna Medium/High/Xhigh/Max: 38/46/49/51 points

That's a massive difference:

- 30 points is Gemma 4 31B territory

- 50 points is GLM-5.2 (744B) territory.

Re: Advancing the price-performance frontier with GPT‑5.6

#127
post #59

Might just resub. Will experiment with Luna next sessions. 5 h window is not working very well for me. But if I can drop down to Luna at 20-30 % left and comfortably ride out the wave then.. that might just work.

They got rid of the 5hr quota, it's just weekly quotas now.

They're supposed to bring 5h today.

Re: Advancing the price-performance frontier with GPT‑5.6

#128
post #106
post #82

Earlier quoted context omitted.

Totally untrue. Luna and Sonnet 5 are very comparable: https://artificialanalysis.ai/#intelligence Luna is an extremely strong model.

> Luna is an extremely strong model. By benchmarks, which sadly is a poor measure. Yes Luna is a good model under certain circumstances. Whether it is great for general usage is another story. Sonnet is definitely better when prompts are more vague and it needs to decide things. Luna generally sticks to things very strictly and goes off in bad ways.

Yes but there is a big, big market for subagents to consume lots of tokens cheaply and condense information up to parent agents. Luna would not be my choice for planning. But an explorer to comb through a codebase to find relevant parts? Or for enterprise retrieval, where it needs to search across many different types of data to see where to focus efforts for a smarter model? Or to wake up periodically to evaluate some conditions and determine if a bigger model should be spun up? Definitely.

I've previously found flash (for all the hate it gets) to be good for these kinds of things. Haiku was fine but it's ancient.

Re: Advancing the price-performance frontier with GPT‑5.6

#129

"Half the money I spend on advertising is wasted; the trouble is I don't know which half." -John Wanamaker This applies even more strongly to model choosing. I know for a fact that majority of my work doesn't require a very strong model, but separating the trivial and non-trivial tasks is a famously hard problem (if at all decidable).

You just need a very strong frontier model to do triage of your tasks.

/s

Re: Advancing the price-performance frontier with GPT‑5.6

#130
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

there is a ton of downward price pressure from Chinese open weight models
Post reply on HN