Live data from Hacker News

Advancing the price-performance frontier with GPT‑5.6

openai.com

151–160 of 424 posts

Re: Advancing the price-performance frontier with GPT‑5.6

#151
post #138

"too cheap to meter" and Luna is still more expensive than deepseek-v4 pro

It's marketing hyperbole, but Luna is more intelligent per dollar than deepseek-v4 pro. Cost means nothing without the associated capability

is it? i don't know how we measure these things, but here's one measurement that says v4 pro is better than luna: https://artificialanalysis.ai/models/comparisons/gpt-5-6-lun...

presumably it's a much bigger model

Re: Advancing the price-performance frontier with GPT‑5.6

#152
post #101
post #58

Earlier quoted context omitted.

lots of places, actually. not everyone wants to be attached to the Silicon Valley culture, and that line alone will guarantee practically any workplace. that person is going to find out what work-life balance is :)

Sure, but those places don’t need such lofty resumes to begin with.

With the right kind of credentials, it's not about need, it's about want. Flip the roles and let yourself become an object of desire, an aspirational hire.

Re: Advancing the price-performance frontier with GPT‑5.6

#153

Earlier quoted context omitted.

https://taalas.com/ has done it already for a wildly obsolete model. 14000 tokens per second. https://chatjimmy.ai/ is their interactive. Tiny context, very dumb, but absurdly fast. Imagine this as a tool call for claude code for trivial changes - the tool call from the harness takes longer than the execution.

Holy crap, I was not prepared for how fast it responded. I just wrote "Just wanted to see how fast you are! Can you write me a quick story about a tiger who lives inside a block of cheese the size of a house?" I pressed Enter, and the response was instant . > Generated in 0.037s • 14,205 tok/s This is unbelievable.

For what its worth the frontier lab models can surely be a lot faster if they wanted them to be but theyre supply constrained so theyre doing stuff like multi tenancy. Since you cant self host them no one outside the labs really knows speed as a solo tenant

Re: Advancing the price-performance frontier with GPT‑5.6

#154
post #100
post #14

> Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, I don't have the words. I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.

> Seeing spikes like this makes me question about where the floor really is. You mean they increased the price and then cut it back and now it is amazing? Luna had a price hike vs mini (its previous replacement). The cut now just puts it back in that ball park. Not that this isn't good news, but what's impressive?

Had to ctrl+f for someone saying this.

I typically do lots of mini calls for research (100s of millions or something in that ball park). Newer models made that absolutely impossible, and the fact that the older ones are starting to get deprecated made me switch to e.g. deepseek for some of my runs. We'll see if I move back after this.

Re: Advancing the price-performance frontier with GPT‑5.6

#155

Earlier quoted context omitted.

Totally possible that humans aren't actually that intelligent.

As well as the existing intelligence being swayed by emotions, hormones, circadian rhythms, stress, peer pressure, propaganda, and survival instincts.

That's a bit like saying a tail is swayed by a dog, as if it could exist without one, or would have anything to do if it did.

Re: Advancing the price-performance frontier with GPT‑5.6

#156

"Half the money I spend on advertising is wasted; the trouble is I don't know which half." -John Wanamaker This applies even more strongly to model choosing. I know for a fact that majority of my work doesn't require a very strong model, but separating the trivial and non-trivial tasks is a famously hard problem (if at all decidable).

You just need a very strong frontier model to do triage of your tasks. /s

That's not necessarily a joke; the article proposes exactly that.

Re: Advancing the price-performance frontier with GPT‑5.6

#157
post #115

Earlier quoted context omitted.

When model intelligence reliably hits 90%-95% of current day knowledge worker tasks, they are going to burn those weight to silicon and we will see another 10X improvement in price/performance frontier. The dynamic GPU clusters will be used for the 5% of tasks, and pushing out the frontier. Also there will be a set of knowledge tasks that are not done today (because they are too difficult for most knowledge workers),…

Burning the weights into silicon would be many orders of magnitude increase, not just 10x. It's kind of crazy that this hockey stick the AI hype bros talk about seems more and more every day like it might be real

And don't underestimate how much money Google, Microsoft, Amazon and Meta still have to spend on this tech.

Blocking Fable for sure made it very politicl a lot sooner than i expected it to happen.

and because China already has massive problems of getting access, they are pushing it on hardware too like what Huawai did without EUV.

It seems China is already able to do DUV a lot sooner than others expected.

Re: Advancing the price-performance frontier with GPT‑5.6

#158
post #2

> GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less Looks like the Chinese models are really making a dent. Having 3 different price categories with the "most affordable" one still costing more than GLM 5.2 never made sense.

It all comes back to electricity cost. China has cheaper electricity so as long as China keeps pace there is no way for American companies to undercut them. Each boolean operation in China is cheaper than the one in America. > China: Household rates average around $0.08 / kWh (¥0.53/kWh). vs > US: Household rates average around $0.16 / kWh, though regional variation is massive—ranging from ~$0.10/kWh in low-cost stat…

I don't know, non of the chinese models I use are served from China. And they are still cheap.

Re: Advancing the price-performance frontier with GPT‑5.6

#159

Earlier quoted context omitted.

https://taalas.com/ has done it already for a wildly obsolete model. 14000 tokens per second. https://chatjimmy.ai/ is their interactive. Tiny context, very dumb, but absurdly fast. Imagine this as a tool call for claude code for trivial changes - the tool call from the harness takes longer than the execution.

Holy crap, I was not prepared for how fast it responded. I just wrote "Just wanted to see how fast you are! Can you write me a quick story about a tiger who lives inside a block of cheese the size of a house?" I pressed Enter, and the response was instant . > Generated in 0.037s • 14,205 tok/s This is unbelievable.

It's crazy. Are they doing any precomputing as you type, I wonder if you paste a block of text is it the same speed.
Post reply on HN