Live data from Hacker News

GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

openrouter.ai

321–330 of 479 posts

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#321
post #182

Earlier quoted context omitted.

Massively subsidized. As soon as my Claude switches from subscription to overage I have to tap out quickly.

This is a big part of the reason I went local-only. Subscription limits are horrible for having a decent workflow.

Yeah, it sure was convenient that there was a RAM pricing crisis right when Apple was making local inference viable. All because of a promise that AI companies will buy more of it... with money they don't yet have, whereas Apple does have lots of money.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#322
post #3

Earlier quoted context omitted.

It's really only between Anthropic and OpenAI for many of my use cases, since I have a Zero Data Retention agreement with both. I'm not trusting random inference providers and especially not Elmo with sensitive data.

You do not have a ZDR with Anthropic. For Mythos and even Fable they require prompt retention on their end. edit: or more precisely if you want to access Mythos/Fable ZDR does not apply, and depending on config the exclusion can affect other models.

I don't have experience directly with Anthropic, but I wouldn't assume anything with such high confidence without knowing the persons situation. No, an individual off the street or with an LLC and 5 employees isn't going to get a special deal from someone like Anthropic.

But if their employer is bringing millions of dollars of potential spend to the table, can tell you from years of experience that turns a lot of 'no's' to 'yes'.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#323
post #278

The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead o…

I shifted from DeepSeek v4 Flash 0731 to Gemini 3.7 flash on openrouter and price shoots up almost double with no visible change in outcome. So, today I reverted back.

Try DS through their own API if that's feasible, AFAIK they're much cheaper than through OS due to cache hit rates.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#324
post #60

Earlier quoted context omitted.

I think Fable's dominance is overstated. It definitely has the lead, but quantifying what that lead actually is is really hard. I'm using GPT 5.6 Sol to do some shit that I personally would consider "crazy" - low level undocumented hardware driver alchemy, reverse engineering highly obfuscated code, even a bit of screwing around with a rendering engine in Vulkan, really just about the most complex tasks I can get any…

AI-pilled obsession with "taste" is bordering on insanity It's just vibes

That's what they're always going to be, so not sure what would be "insane" about it. They literally feed on and emit natural language, and are put to work on informally defined, arbitrary tasks.

When people figure out any reliable strategies to test and benchmark them, that's insane, and in the positive sense. This very same issue has been a thing for humans as well forever, and remains only very questionably solved (IQ, academic tests). This is not easy.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#325
post #304
post #302

dang, to avoid confusion from the title perhaps this should be edited to: “OpenRouter temporarily cutting GPT-5.6 Sol pricing by 50%”

That suggests this is being funded by OpenRouter, and there's no indication this is the case (and I doubt OpenRouter can afford it; why would they anyway).

There also doesn’t seem to be an official indication that this is funded by OpenAI, hence the confusion in the thread.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#326
post #314
post #308

Earlier quoted context omitted.

Long-context agentic tasks and Rust engineering are our use cases where Kimi definitely is better than Sol. We can measure our own systems and the numbers say that Sol has no chance against K3 or Opus, and K3 is so so so much cheaper than Opus right now. You cannot just look at the price tags for these models, you must eval and see the price per task. In our previous eval rounds Sol was more expensive than Opus (with…

China's 50 Cent Party being a real and noticeable thing (and the two biggest things they like to shill is open weight Chinese models and the futility of resisting a Taiwan invasion), I have to take things like this with a healthy dose of skepticism without corroborating data, since independent evals didn't show the price per task lead you're showing. If there's independent data showing this feel free to share a link,…

Internal reports from company? Maybe not. I'm just saying you have to eval eval eval if you are working in this industry. There's a ton of victories in price, and price is right now the key thing all the customers are talking about.

It's not always Chinese models. For example GLM 5.2 just did not work for us at all. And Gemini is still the best cheap model for non-text agents.

If you don't have a good eval set and if you don't check the models weekly, you are missing on things. And Opus 4.8 is still the absolute quality king for agentic tasks. Too bad it's so expensive.

And the clearest thing here is that Fable, Opus, and Sol are all too expensive. I'd say a healthy 75% cut to token prices and they are back in competition.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#327
post #201

Where is the official source for this? OpenAI's docs still show non-discounted pricing https://developers.openai.com/api/docs/models/gpt-5.6-sol

It's discounted if used via OpenRouter, not the official API.

And it's not just on OpenRouter; Vercel offers the same:

https://vercel.com/changelog/gpt-5-6-sol-is-50-off-on-ai-gat...

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#328
post #278

The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead o…

It's funny how quickly we went from "the greedy US companies are subsidizing prices to keep competitors out of the market" to "the greedy US companies are overcharging because they are greedy."

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#329
OpenAI is making some really boneheaded moves these days.

They're reacting instead of leading, basically.

Cutting API prices 50% while millions of your paying subscribers have had their limits slashed and are all literally looking at the salivatingly-cheap chinese API prices availalbe on openrouter...

Not only did OpenAI and all of their cash somehow MISS the opportunity to purchase OpenRouter ...

Now they're giving a discount on an API that nobody even uses (get real, nobody's paying API prices to OpenAI ...

I calculated a 5.6 sol coding session the other day ... $680+ USD ... and it actually destroyed the codebase it was working on during that session).

Needless to say, I will not be spending another dime with Codex or OpenAI.

This entire Codex reset limit fiasco has taught me they are not to be trusted.

Deepseek, here I come.

Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

#330
post #278

The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead o…

It's funny how quickly we went from "the greedy US companies are subsidizing prices to keep competitors out of the market" to "the greedy US companies are overcharging because they are greedy."

They are subsidizing the non-API use cases and overcharging on the API use cases.
Post reply on HN