But what no one mentions is that the price is going from a starting point of $0.16 to $0.60, so basically they're charging nearly four times as much.
Like _aavaa_ said, make sure you're comparing the right vals 1:1. There's different costs for cache hit, cache misses, output tokens, etc. This one seems, during non-peak hours, cheaper. Peak hours are obviously more expensive, if they're gonna be 2x non-peak pricing. But, that might end up decreasing in the future.
DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
71–80 of 221 posts
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#72hope deepseek makes me change my setup again
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#73But what no one mentions is that the price is going from a starting point of $0.16 to $0.60, so basically they're charging nearly four times as much.
Like _aavaa_ said, make sure you're comparing the right vals 1:1. There's different costs for cache hit, cache misses, output tokens, etc. This one seems, during non-peak hours, cheaper. Peak hours are obviously more expensive, if they're gonna be 2x non-peak pricing. But, that might end up decreasing in the future.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#74Beta testers report >400 TPS. https://www.geeky-gadgets.com/deepseek-v4-1-flash-review/ I hope some of those speed increases will make it to production.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#75Sounds nice! But, the web ui chat version of flash has very poor language following abilities in my experience: You may ask it something in English, and get a thinking chain in Chinese with an answer in Chinese, or an English thinking chain and an English answer. Using the retry button on the same question has a 50/50 chance of any of those results. Sometimes, asking something in English, but where information are mo…
It's not just web chat, V4 Flash 7/31 suffers from a lot of pathological behavior in coding harnesses as well, e.g. infinite loops, hallucinations, premature termination, and invalid tool calls.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#76> all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price If I'd carefully tested and optimized prompts against Pro I wouldn't be keen on this particular news. I feel like API model providers should lean towards not swapping out models on their paying customers, no matter how much "better" the new model is meant to be.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#77I will continue to be amazed by how much power you get from DeepSeek Flash for the cost. I have let that puppy lose on so many projects and it is has never let me down. It can build and entire Rails app in no time and even do the tests. For most things, I don't get why people pay the money for Claude. DeepSeek Flash is my default agent in Omarchy.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#78It's interesting that this is the third lab to find problems with larger models. Earlier last year oAI was rumoured to have failed their large pretrain. Now google has problems with their pro series, and ds just announced the same. There are some rumours on chinese forums talking about problems with the pretraining phase, so this is not mid/post training related. I wonder if this comes from using the bad architecture…
The announcement specifically says 4.1 Pro will be released in the future.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#79It's interesting that this is the third lab to find problems with larger models. Earlier last year oAI was rumoured to have failed their large pretrain. Now google has problems with their pro series, and ds just announced the same. There are some rumours on chinese forums talking about problems with the pretraining phase, so this is not mid/post training related. I wonder if this comes from using the bad architecture…
V4 flash and V4 pro feel very similar, which would make sense if they were pre-trained on largely the same corpus.
All that would suggest to me is that V4 Flash is capable of absorbing the data they’re throwing at it, and we’re still nowhere near the data limits of their larger 1.6T model
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#80if this beats GLM 5.3 flash, I am sold