Advancing the price-performance frontier with GPT‑5.6
11–20 of 424 posts
Re: Advancing the price-performance frontier with GPT‑5.6
#12> GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less Looks like the Chinese models are really making a dent. Having 3 different price categories with the "most affordable" one still costing more than GLM 5.2 never made sense.
> China: Household rates average around $0.08 / kWh (¥0.53/kWh).
vs
> US: Household rates average around $0.16 / kWh, though regional variation is massive—ranging from ~$0.10/kWh in low-cost states (like Washington or Louisiana) to $0.30–$0.45+/kWh in high-cost areas like California or Hawaii.
Re: Advancing the price-performance frontier with GPT‑5.6
#13Re: Advancing the price-performance frontier with GPT‑5.6
#14I don't have the words.
I genuinely thought we were in a stage where we were plateauing and going in for 5-10% improvements over months. Seeing spikes like this makes me question about where the floor really is.
Re: Advancing the price-performance frontier with GPT‑5.6
#15> The kernel work helped reduce the end-to-end cost of serving the model by 20%, while its experiments increased token-generation efficiency by more than 15%. If the cost of serving GPT-5.6 just dropped by 20%, does that add up to literally billions of dollars in savings per month? We know Anthropic spend $1.25 billion renting inference capacity from SpaceX (in two Colossus datacenters) from the SpaceX IPO, but we do…
imagine writing that on your resume > reduced inference cost by 20 percent saving company x billion dollars per month
I have no doubt that further work was required to enable this, but it's still very cool to be possible to say that.
Re: Advancing the price-performance frontier with GPT‑5.6
#16Model segmentation & distillation like this that asks the consumers to pick exactly which version of the algorithm will solve their problem is evidence for lack of intelligence instead of its presence.
Re: Advancing the price-performance frontier with GPT‑5.6
#17> The kernel work helped reduce the end-to-end cost of serving the model by 20%, while its experiments increased token-generation efficiency by more than 15%. If the cost of serving GPT-5.6 just dropped by 20%, does that add up to literally billions of dollars in savings per month? We know Anthropic spend $1.25 billion renting inference capacity from SpaceX (in two Colossus datacenters) from the SpaceX IPO, but we do…
Re: Advancing the price-performance frontier with GPT‑5.6
#18> GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less Looks like the Chinese models are really making a dent. Having 3 different price categories with the "most affordable" one still costing more than GLM 5.2 never made sense.
Re: Advancing the price-performance frontier with GPT‑5.6
#19Re: Advancing the price-performance frontier with GPT‑5.6
#20Not sure who would use Terra anymore. Pair Luna High/Xhigh with Sol Medium and that's your power stack