About 3x increase. Luna is now a much better deal. Hope they don't increase their prices in response.
Luna has one vendor and they can change the price any time. Deepseek is open and has dozens of vendors competing to serve it.
Thank you in advance!
41–50 of 201 posts
About 3x increase. Luna is now a much better deal. Hope they don't increase their prices in response.
Luna has one vendor and they can change the price any time. Deepseek is open and has dozens of vendors competing to serve it.
Thank you in advance!
I pay for Google AI Pro (Bought a year in advance) and Gemini is so bad, I burned through 75% of my five hour allowance trying to get it to fix something. I pasted the same prompt into OpenCode, set to Deepseek v4 flash free and did it first try. I'm was going to purchase Opencode GO to try it, but seems my timing is really bad :( hope it doesn't go up too much in Opencode or they find other providers. Bad timing!
Old DeepSeek Flash 0731 prices have been independently reproduced.[1] The issue is DeepSeek being inundated and not having capacity to serve the demand, hence the price increases to significantly dampen demand. Never mind international demand either--just think about the magnitude of Chinese domestic demand. Prices for anything related to AI or computing in general (mobile phones, cloud data centre hosting, etc) will…
Reproduced on the CUDA stack right? Let's say DeepSeek is being forced to use the CANN stack, and the new pricing reflects the cost when 100% of inference is done with Huawei chips. Then, I suppose we can infer that: * CANN stack is 1.5x~2.3x less efficient in compute * CANN stack has 6x lower inter-connect capacity > computer chips once again become a commodity Ascend 950 is going for $7k to $9k with mediocre lookin…
The CEO of DeepSeek recently revealed to investors a lot about the resources available to DeepSeek, and the gap between Huawei and NVIDIA. Select quotes from the transcript (translation is a bit patchy on the source website though):
"We currently have roughly 20,000 H-equivalent compute cards"
"Huawei 950—right now Huawei gives us 16,000 cards, this should be publicly stateable."
"Like Huawei gives us roughly 16,000 cards of capacity, internet giants maybe get a hundred-something thousand, we get ten-something thousand—I think this ratio is also relatively... but this is probably just how much capacity Huawei has."
"16,000 Huawei 950 cards only equal 4,000 B-series cards."
"Huawei’s supernode, Huawei’s 950 supernode, in performance and price can completely substitute for NVIDIA’s GB200, GB300. The price is definitely more expensive, but limitedly so. Fifty percent more expensive, a hundred percent more expensive—a hundred percent more doesn’t matter, two hundred percent more doesn’t matter. For example, a hundred percent more expensive—I think it can already be considered a price-level substitute."
"I think domestic hardware might need a few years."
"I don’t quite believe that five years from now, we’ll still be stuck on the production capacity problem. Right now we’re definitely stuck on the production capacity problem—this year, next year, the year after, I think we might still be stuck on the production capacity problem, but five years later, I think maybe not necessarily—I’m still relatively optimistic."
[1] https://www.fredgao.com/p/deepseeks-liang-wenfeng-breaks-his
Earlier quoted context omitted.
Luna vs v4 Flash, sure. But DeepSeek v4 Pro is a far more capable model and still cheaper than anything that it competes with, from what I can see.
Luna is multimodal.
Earlier quoted context omitted.
Luna has one vendor and they can change the price any time. Deepseek is open and has dozens of vendors competing to serve it.
Can you help me find a few names? These are inference providers running open weights models, right? Thank you in advance!
So OpenAI cut Luna's price by 5x, DeepSeek increased price by 5x! If I'm reading the benchmarks right, they now went from being much cheaper than Luna (but twice as slow), to being roughly same price (but twice as slow). So all else being equal, where I would previously have used DeepSeek, I can just use Luna, and get the same result twice as fast? (Yeah I know benchmarks are mostly nonsense, but the ones measuring t…
So OpenAI cut Luna's price by 5x, DeepSeek increased price by 5x! If I'm reading the benchmarks right, they now went from being much cheaper than Luna (but twice as slow), to being roughly same price (but twice as slow). So all else being equal, where I would previously have used DeepSeek, I can just use Luna, and get the same result twice as fast? (Yeah I know benchmarks are mostly nonsense, but the ones measuring t…
Also if you're considering Luna, I assume you don't care about this but I think it's worth pointing out: a major advantage of DS is the ability to self-host or choose a different host. As a customer that gives you much more negotiating power and potential privacy guarantees.
Earlier quoted context omitted.
> EDIT: formatting Keep at it, I believe in you.
My apologies to any mobile users, but for the desktop folk: Provider, Model Billing Input Output Cache read Cache write DeepSeek V4-Flash Old $0.1400 $0.2800 $0.0028 - V4-Flash New Off-Peak $0.2200 (1.6x) $0.6600 (2.4x) $0.0070 (2.5x) - V4-Flash New Peak $0.4400 (3.1x) $1.3200 (4.7x) $0.0140 (5.0x) - V4-Pro Old $0.4350 $0.8700 $0.0036 - V4-Pro New Off-Peak $0.6600 (1.5x) $1.9800 (2.3x) $0.0220 (6.1x) - V4-Pro New Pea…
Earlier quoted context omitted.
My apologies to any mobile users, but for the desktop folk: Provider, Model Billing Input Output Cache read Cache write DeepSeek V4-Flash Old $0.1400 $0.2800 $0.0028 - V4-Flash New Off-Peak $0.2200 (1.6x) $0.6600 (2.4x) $0.0070 (2.5x) - V4-Flash New Peak $0.4400 (3.1x) $1.3200 (4.7x) $0.0140 (5.0x) - V4-Pro Old $0.4350 $0.8700 $0.0036 - V4-Pro New Off-Peak $0.6600 (1.5x) $1.9800 (2.3x) $0.0220 (6.1x) - V4-Pro New Pea…
Seems to still be cheaper at worst case scenario on a pay as you schedule. The price increase isn't ideal, but still seems like a good deal to me.
Earlier quoted context omitted.
My apologies to any mobile users, but for the desktop folk: Provider, Model Billing Input Output Cache read Cache write DeepSeek V4-Flash Old $0.1400 $0.2800 $0.0028 - V4-Flash New Off-Peak $0.2200 (1.6x) $0.6600 (2.4x) $0.0070 (2.5x) - V4-Flash New Peak $0.4400 (3.1x) $1.3200 (4.7x) $0.0140 (5.0x) - V4-Pro Old $0.4350 $0.8700 $0.0036 - V4-Pro New Off-Peak $0.6600 (1.5x) $1.9800 (2.3x) $0.0220 (6.1x) - V4-Pro New Pea…
Seems to still be cheaper at worst case scenario on a pay as you schedule. The price increase isn't ideal, but still seems like a good deal to me.
I’m curious about how openrouter and Luna prices will change in response.