Live data from Hacker News

DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

news.ycombinator.com

131–140 of 214 posts

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#132
post #15

Since a few months, I almost exclusively use the Chinese "flash" models for my needs. They are a joy and they cost pennies per answer. Great job.

I am legitimately more excited for this release than any frontier models at this point. I don't need a model that can invent new mathematics. I need something that is fast, cheap, and consistent. Give me that and I can build and scale.

LLMs are not consistent

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#133
post #114
post #20

Sounds nice! But, the web ui chat version of flash has very poor language following abilities in my experience: You may ask it something in English, and get a thinking chain in Chinese with an answer in Chinese, or an English thinking chain and an English answer. Using the retry button on the same question has a 50/50 chance of any of those results. Sometimes, asking something in English, but where information are mo…

All flash llms have this problems. gemini. I start to a new chat write in german and suddenly it answers in english. I take the free chat gpt one writ with it in polish suddenly english.

You see this on Reddit where the bot accounts will just comment in German, French or Italian randomly (and other bot accounts responding to it won't even bat an eye, responding in English as if it's the most natural thing in the world)

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#134
post #67

>In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price Please don't do this kind of thing. If a user has validated a workflow on V4 Pro, they might not want to suddenly start testing it in production on V4.1 Flash. Instead, keep V4 Pro around but dep…

In this case, Deepseek organization is under a lot of pressure due to compute constraints. It would be better if they just throw a 404 instead of rerouting though so customers are not surprised by subtle changes in behavior.

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#135
post #20

Sounds nice! But, the web ui chat version of flash has very poor language following abilities in my experience: You may ask it something in English, and get a thinking chain in Chinese with an answer in Chinese, or an English thinking chain and an English answer. Using the retry button on the same question has a 50/50 chance of any of those results. Sometimes, asking something in English, but where information are mo…

Yep. I faced the exact same issue. Too many times. And then just gave up.

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#136

It's interesting that this is the third lab to find problems with larger models. Earlier last year oAI was rumoured to have failed their large pretrain. Now google has problems with their pro series, and ds just announced the same. There are some rumours on chinese forums talking about problems with the pretraining phase, so this is not mid/post training related. I wonder if this comes from using the bad architecture…

Where are you seeing them having an issue with the larger (Pro) model? The announcement specifically says 4.1 Pro will be released in the future.

A month ago new V4 Flash 0731 checkpoint was better than existing V4 Pro. They've kept serving Pro, it was updated 13 days later (0813 checkpoint).

Now, 4 weeks later new Flash checkpoint (0910?) is again better than existing Pro. Same situation, but Pro is taken offline this time.

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#137
post #107

Earlier quoted context omitted.

So you did not do what the post you replied to suggested?

Oh, I’ve tried that too. It will promise to keep it in English from here on out, then switch back to Chinese after two or three exchanges. When ever it needs to do a web search, it seems to load so much Chinese text that it forgets any language instructions. Just thought my experience yesterday was more to the point. Right now the chat is absolutely hopeless.

huh. mine only does this on the first turn on a new computer. Once I've told it once it seems to be entirely sticky on that device from then on

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#138
post #114

Earlier quoted context omitted.

All flash llms have this problems. gemini. I start to a new chat write in german and suddenly it answers in english. I take the free chat gpt one writ with it in polish suddenly english.

You see this on Reddit where the bot accounts will just comment in German, French or Italian randomly (and other bot accounts responding to it won't even bat an eye, responding in English as if it's the most natural thing in the world)

It's a bit more complicated than that because Reddit now automatically tries to translate comments not in the user's language.

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#139
post #67

>In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price Please don't do this kind of thing. If a user has validated a workflow on V4 Pro, they might not want to suddenly start testing it in production on V4.1 Flash. Instead, keep V4 Pro around but dep…

LLMs add enough nondeterminism to a workflow. Swapping them without the user knowing adds substantially more.

Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

#140
post #65
post #32

Earlier quoted context omitted.

What are you talking about? Current flash prices are 0.66 for output, this is dropping it to 0.60.

This is the notice from DeepSeek regarding their API: We will adjust the pricing for the Flash series effective from 12:00 Beijing Time on September 10, 2026. During off-peak hours, the unit price will be $0.003 for input cache hits, $0.15 for input cache misses, and $0.6 for output. Peak-hour prices will be double the off-peak rates. Please plan your usage accordingly. -----------------------------------------------…

> Hoje em sites como openrouter o valor é de $0.16 output .

You're comparing different providers then. DeepSeek price on OpenRouter is $0.66 output.

Post reply on HN