Earlier quoted context omitted.
Massively subsidized. As soon as my Claude switches from subscription to overage I have to tap out quickly.
This is a big part of the reason I went local-only. Subscription limits are horrible for having a decent workflow.
GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
321–330 of 479 posts
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#322Earlier quoted context omitted.
It's really only between Anthropic and OpenAI for many of my use cases, since I have a Zero Data Retention agreement with both. I'm not trusting random inference providers and especially not Elmo with sensitive data.
You do not have a ZDR with Anthropic. For Mythos and even Fable they require prompt retention on their end. edit: or more precisely if you want to access Mythos/Fable ZDR does not apply, and depending on config the exclusion can affect other models.
But if their employer is bringing millions of dollars of potential spend to the table, can tell you from years of experience that turns a lot of 'no's' to 'yes'.
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#323The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead o…
I shifted from DeepSeek v4 Flash 0731 to Gemini 3.7 flash on openrouter and price shoots up almost double with no visible change in outcome. So, today I reverted back.
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#324Earlier quoted context omitted.
I think Fable's dominance is overstated. It definitely has the lead, but quantifying what that lead actually is is really hard. I'm using GPT 5.6 Sol to do some shit that I personally would consider "crazy" - low level undocumented hardware driver alchemy, reverse engineering highly obfuscated code, even a bit of screwing around with a rendering engine in Vulkan, really just about the most complex tasks I can get any…
AI-pilled obsession with "taste" is bordering on insanity It's just vibes
When people figure out any reliable strategies to test and benchmark them, that's insane, and in the positive sense. This very same issue has been a thing for humans as well forever, and remains only very questionably solved (IQ, academic tests). This is not easy.
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#325dang, to avoid confusion from the title perhaps this should be edited to: “OpenRouter temporarily cutting GPT-5.6 Sol pricing by 50%”
That suggests this is being funded by OpenRouter, and there's no indication this is the case (and I doubt OpenRouter can afford it; why would they anyway).
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#326Earlier quoted context omitted.
Long-context agentic tasks and Rust engineering are our use cases where Kimi definitely is better than Sol. We can measure our own systems and the numbers say that Sol has no chance against K3 or Opus, and K3 is so so so much cheaper than Opus right now. You cannot just look at the price tags for these models, you must eval and see the price per task. In our previous eval rounds Sol was more expensive than Opus (with…
China's 50 Cent Party being a real and noticeable thing (and the two biggest things they like to shill is open weight Chinese models and the futility of resisting a Taiwan invasion), I have to take things like this with a healthy dose of skepticism without corroborating data, since independent evals didn't show the price per task lead you're showing. If there's independent data showing this feel free to share a link,…
It's not always Chinese models. For example GLM 5.2 just did not work for us at all. And Gemini is still the best cheap model for non-text agents.
If you don't have a good eval set and if you don't check the models weekly, you are missing on things. And Opus 4.8 is still the absolute quality king for agentic tasks. Too bad it's so expensive.
And the clearest thing here is that Fable, Opus, and Sol are all too expensive. I'd say a healthy 75% cut to token prices and they are back in competition.
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#327Where is the official source for this? OpenAI's docs still show non-discounted pricing https://developers.openai.com/api/docs/models/gpt-5.6-sol
It's discounted if used via OpenRouter, not the official API.
https://vercel.com/changelog/gpt-5-6-sol-is-50-off-on-ai-gat...
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#328The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead o…
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#329They're reacting instead of leading, basically.
Cutting API prices 50% while millions of your paying subscribers have had their limits slashed and are all literally looking at the salivatingly-cheap chinese API prices availalbe on openrouter...
Not only did OpenAI and all of their cash somehow MISS the opportunity to purchase OpenRouter ...
Now they're giving a discount on an API that nobody even uses (get real, nobody's paying API prices to OpenAI ...
I calculated a 5.6 sol coding session the other day ... $680+ USD ... and it actually destroyed the codebase it was working on during that session).
Needless to say, I will not be spending another dime with Codex or OpenAI.
This entire Codex reset limit fiasco has taught me they are not to be trusted.
Deepseek, here I come.
Re: GPT-5.6 Sol Pricing Cut by 50% on OpenRouter
#330The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead o…
It's funny how quickly we went from "the greedy US companies are subsidizing prices to keep competitors out of the market" to "the greedy US companies are overcharging because they are greedy."