Live data from Hacker News

Advancing the price-performance frontier with GPT‑5.6

openai.com

361–370 of 424 posts

Re: Advancing the price-performance frontier with GPT‑5.6

#362
post #239

Earlier quoted context omitted.

This has already been updated with the new prices?

The intelligence scores are absolute, not per $. DeepSeek Pro is more capable than Luna regardless of the cost.

There's also a cost graph on the same page and the context of this conversation is cost vs performance frontier.

Re: Advancing the price-performance frontier with GPT‑5.6

#364
post #115

Earlier quoted context omitted.

Burning the weights into silicon would be many orders of magnitude increase, not just 10x. It's kind of crazy that this hockey stick the AI hype bros talk about seems more and more every day like it might be real

https://taalas.com/ has done it already for a wildly obsolete model. 14000 tokens per second. https://chatjimmy.ai/ is their interactive. Tiny context, very dumb, but absurdly fast. Imagine this as a tool call for claude code for trivial changes - the tool call from the harness takes longer than the execution.

Answers like 8Bish model it appears, but instant response

Re: Advancing the price-performance frontier with GPT‑5.6

#365
post #302
post #159

Earlier quoted context omitted.

It's crazy. Are they doing any precomputing as you type, I wonder if you paste a block of text is it the same speed.

The magic is in the fact that they essentially have an ASIC llm device. There is no other trickery. The problem they will face is that it is actually locked in silicon, so upgrading models will be difficult, and likely require new hardware each time.

Is the size of the ASIC limiting factor or can they infinitely tensor parallelize? If thats possible then it would make it only a matter of economy of scale, there are good enough models already for people to invest in that kind of platform

Re: Advancing the price-performance frontier with GPT‑5.6

#367
post #361
post #351

Earlier quoted context omitted.

Source?

https://www.wheresyoured.at/exclusive-openai-financials/

The company losing money does not mean the model inference in API is 70% subsidized- that’s a crazy leap in logic. Obviously the massive number of free chatgpt users are getting subsidized 100%.

Re: Advancing the price-performance frontier with GPT‑5.6

#368
post #363

GPT-5.6 Luna xhigh/max from Codex is my daily driver now. Good economics, good performance. It sits in the most attractive quadrant on the Intelligence Index.

any particular reason not using terra or sol, given the generous limit quotas and frequent codex resets?

Re: Advancing the price-performance frontier with GPT‑5.6

#369
post #286

Earlier quoted context omitted.

That's surprising. Does it have no sub agent support at all or does it just use the same agent as the parent?

They do have subagents, released v2 of that feature with the launch of 5.6 model series in fact. It's just... very poorly executed, is a significant regression from subagents v1 and thousands of miles behind subagents of Claude code. - Models that can be launched as subagents are hardcoded (can only be another Sol or Terra, but not Luna). Most of the time it'll just launch same model as parent anyway. - They encrypt…

Which is why I use CLI subprocesses as subagents... Codex in 2026 is still nowhere near where CC was in summer of 2025...

Re: Advancing the price-performance frontier with GPT‑5.6

#370

How is that economically possible? I’m so confused by those prices

how many times do you have to be metaphorically hit in the head with a brick before you realize inference margins at api pricing were 80%+

Uh… are you okay? That’s a completely disproportionate, aggressive tone. How is that warranted
Post reply on HN