DeepSeek V4 Flash 0731
121–130 of 479 posts
Re: DeepSeek V4 Flash 0731
#122Re: DeepSeek V4 Flash 0731
#123I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…
How is $5/day irrelevant? In the $150/mo range you can get effectively unlimited usage of GPT 5.6 Sol (Pro plan). Why use a much weaker model for the same price?
i used for work where i did less and it quickly reaches thousands if you're not careful. i can already see what some will say: skill issue et cetera - whatever.
Re: DeepSeek V4 Flash 0731
#124This latest DeepSeek is almost at the "too cheap to meter" level. That's going to be a larger unlock than models like Fable/Mythos that are way too expensive to justify, IMO. What secret sauce do they have?
Re: DeepSeek V4 Flash 0731
#125Kimi K3 was an interesting model only a month ago, and now we're looking at the same performance for 1/20th of the price. Wild how fast this is advancing.
Not for long, Deepseek is saying they will have a significant price jump soon. They really shouldn’t do it because they are on the cusp of capturing the scalable API market.
Re: DeepSeek V4 Flash 0731
#126Re: DeepSeek V4 Flash 0731
#127Earlier quoted context omitted.
Why? It's open weight, there are plenty providers on open router that are serving the latest v4 flash at 0.14/0.28 $.
Yes, but even the cheapest providers on OpenRouter are charging at least 10x what DeepSeek does for cached input tokens, which is where DeepSeek gets most of the cheapness.
Re: DeepSeek V4 Flash 0731
#128It's serviceable but, like many Chinese models, it uses a lot of tokens to get work done.
That's how I handle the Qwen27B and 35B
Re: DeepSeek V4 Flash 0731
#129It's serviceable but, like many Chinese models, it uses a lot of tokens to get work done.
It felt like a rocket compared to GLM 5.2 though. Are Chinese models generally token-heavy?
Re: DeepSeek V4 Flash 0731
#130I strongly recommend trying this for programming tasks. It is strong (not Fable strong though) with a much better “persona” than Opus, and very different blindspots. If you flip between Claude and this you will find both catch the mistakes of the other before they get out of control. On balance I actually prefer DeepSeek for programming now, because of the way it talks.
[flagged]