Live data from Hacker News

DeepSeek V4 Flash 0731

arcprize.org

121–130 of 479 posts

Re: DeepSeek V4 Flash 0731

#122
post #101
post #93

Earlier quoted context omitted.

These posts have to be Chinese bots, these models are all trash. Used it via OpenCode for an hour, cost me one hour of my life. It is for anything complete trash.

You're mad.

Point 1 finger out, and you point 4 back.

Re: DeepSeek V4 Flash 0731

#123
post #118

I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…

How is $5/day irrelevant? In the $150/mo range you can get effectively unlimited usage of GPT 5.6 Sol (Pro plan). Why use a much weaker model for the same price?

not at all true. if you're truly using it across the board for smaller things (translation of pages, filtering of every individual tweet based on its relevance to you etc), the costs ramp up super quickly.

i used for work where i did less and it quickly reaches thousands if you're not careful. i can already see what some will say: skill issue et cetera - whatever.

Re: DeepSeek V4 Flash 0731

#124

This latest DeepSeek is almost at the "too cheap to meter" level. That's going to be a larger unlock than models like Fable/Mythos that are way too expensive to justify, IMO. What secret sauce do they have?

No secrets—all published. Very efficient attention. Excellent kernels. Great caching subsystem. Small and well trained model.

Re: DeepSeek V4 Flash 0731

#125
post #4

Kimi K3 was an interesting model only a month ago, and now we're looking at the same performance for 1/20th of the price. Wild how fast this is advancing.

Not for long, Deepseek is saying they will have a significant price jump soon. They really shouldn’t do it because they are on the cusp of capturing the scalable API market.

They need to be able to serve their market. The price increase is partly load shedding. If they improve their ability to serve their load, they can always drop it again, as OpenAI did with Luna recently.

Re: DeepSeek V4 Flash 0731

#127

Earlier quoted context omitted.

Why? It's open weight, there are plenty providers on open router that are serving the latest v4 flash at 0.14/0.28 $.

Yes, but even the cheapest providers on OpenRouter are charging at least 10x what DeepSeek does for cached input tokens, which is where DeepSeek gets most of the cheapness.

[deleted]

Re: DeepSeek V4 Flash 0731

#128
post #17

It's serviceable but, like many Chinese models, it uses a lot of tokens to get work done.

If I had the GPU size, hook it up to llama.cpp and setup the --reasoning-budget and reasoning-message; Most of that additional reasoning is a lot of garbage and you can redirect it to useful output.

That's how I handle the Qwen27B and 35B

Re: DeepSeek V4 Flash 0731

#129
post #17

It's serviceable but, like many Chinese models, it uses a lot of tokens to get work done.

It felt like a rocket compared to GLM 5.2 though. Are Chinese models generally token-heavy?

https://artificialanalysis.ai/?cost=intelligence-vs-cost-per...

Re: DeepSeek V4 Flash 0731

#130
post #97
post #44

I strongly recommend trying this for programming tasks. It is strong (not Fable strong though) with a much better “persona” than Opus, and very different blindspots. If you flip between Claude and this you will find both catch the mistakes of the other before they get out of control. On balance I actually prefer DeepSeek for programming now, because of the way it talks.

[flagged]

[deleted]
Post reply on HN