If they can keep up this cadence of Flash leap-frogging the previous Pro, we're in for a good time
Anthropic/OpenAI might step up their anti-distillation defences though.
DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
81–90 of 219 posts
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#82While V4.1 Flash performance and cost looks promising this auto re-routing sounds concerning
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#83Earlier quoted context omitted.
Like _aavaa_ said, make sure you're comparing the right vals 1:1. There's different costs for cache hit, cache misses, output tokens, etc. This one seems, during non-peak hours, cheaper. Peak hours are obviously more expensive, if they're gonna be 2x non-peak pricing. But, that might end up decreasing in the future.
Deepseek already has 2x peak pricing. This is just going to be cheaper across the board.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#84So, will it have vision? (based on deepseek-v4-flash-vision-exp ?)
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#85>In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price Please don't do this kind of thing. If a user has validated a workflow on V4 Pro, they might not want to suddenly start testing it in production on V4.1 Flash. Instead, keep V4 Pro around but dep…
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#86Earlier quoted context omitted.
What are you talking about? Current flash prices are 0.66 for output, this is dropping it to 0.60.
This is the notice from DeepSeek regarding their API: We will adjust the pricing for the Flash series effective from 12:00 Beijing Time on September 10, 2026. During off-peak hours, the unit price will be $0.003 for input cache hits, $0.15 for input cache misses, and $0.6 for output. Peak-hour prices will be double the off-peak rates. Please plan your usage accordingly. -----------------------------------------------…
0.007 -> 0.003
0.22 -> 0.15
0.66 -> 0.60
Each one is now cheaper.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#87Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#88Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#89Earlier quoted context omitted.
What are you talking about? Current flash prices are 0.66 for output, this is dropping it to 0.60.
This is the notice from DeepSeek regarding their API: We will adjust the pricing for the Flash series effective from 12:00 Beijing Time on September 10, 2026. During off-peak hours, the unit price will be $0.003 for input cache hits, $0.15 for input cache misses, and $0.6 for output. Peak-hour prices will be double the off-peak rates. Please plan your usage accordingly. -----------------------------------------------…
Input cache hits (per 1m tokens) - $0.003 Vs. $0.022 Vs. $0.007
Input cache miss (per 1m tokens) - $0.15 Vs. $0.66 Vs. $0.22
Output (per 1m tokens) - $0.6 Vs. $1.98 Vs. $0.66
This is taken from https://api-docs.deepseek.com/quick_start/pricing, and it's comparing only off-peak hours pricing. It looks like V4.1 Flash is cheaper than the current 0731 flash model, and much cheaper than the current V4 Pro model.