Live data from Hacker News

DeepSeek V4 Flash 0731

arcprize.org

21–30 of 474 posts

Re: DeepSeek V4 Flash 0731

#21
post #15

I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…

But DeepSeek now has a warning they’re going to sharply increase their API pricing sometime in the future.

Dax (from Opencode) has tweeted that they can replicate or beat the price with rented GPUs. Deepseeks secret sauce is the incredibly cheap caching (magnitude cheaper than other providers).

vLLM has recently released a similar approach. It's not as effective as what DeepSeek does but still an interesting development.

I have no doubt that in due time other providers will match or perhaps even beat the current DeepSeek prices.

Re: DeepSeek V4 Flash 0731

#22

I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…

>even if it's not SOTA

And, probably 99.99% of people using LLM probably don't even need SOTA anyway.

Re: DeepSeek V4 Flash 0731

#23

Earlier quoted context omitted.

Since the x-axis is log-scaled, DeepSeek is much cheaper than visually implied (mousing over the raw values, it's 1/4th the cost of Luna).

Is this pricing from Deepseek with training on usage?

Per the announcement tweet, BaseTen was the inference provider which has 20% cache cost that is typical: https://www.baseten.co/library/deepseek-v4-flash-0731/

Re: DeepSeek V4 Flash 0731

#25
post #19
post #15

Earlier quoted context omitted.

But DeepSeek now has a warning they’re going to sharply increase their API pricing sometime in the future.

Source?

If you're on the DeepSeek Platform, you'd see this:

"We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice."

Re: DeepSeek V4 Flash 0731

#26
post #15

I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…

But DeepSeek now has a warning they’re going to sharply increase their API pricing sometime in the future.

[deleted]

Re: DeepSeek V4 Flash 0731

#27
They did recently announce they're increasing prices though (got a mail yesterday I think), so not sure this analysis showing it as price outlier will last

Re: DeepSeek V4 Flash 0731

#28
post #7

Price is confounded by VC subsidies, economies of scale, and inference optimizations. I think a more interesting chart would be ARC AGI vs forwards pass flops or ARC AGI vs training tokens. Of course we don't have those numbers for the closed source models or even some of the open weight ones.

DeepSeek V4 Flash 0731 is an open-weights model which means price is determined by competition/invisible hand of the marketplace: https://openrouter.ai/deepseek/deepseek-v4-flash-0731 With the exception of cache costs, all providers have similar input/output costs.

Not counting the cost of making the model, which is subsidized by… someone? The chinese gov i think?

Re: DeepSeek V4 Flash 0731

#29
This latest DeepSeek is almost at the "too cheap to meter" level. That's going to be a larger unlock than models like Fable/Mythos that are way too expensive to justify, IMO.

What secret sauce do they have?

Re: DeepSeek V4 Flash 0731

#30
post #27

They did recently announce they're increasing prices though (got a mail yesterday I think), so not sure this analysis showing it as price outlier will last

That is only when using the DeepSeek API directly. OpenRouter has 24 different providers serving it at existing prices.
Post reply on HN