Since a few months, I almost exclusively use the Chinese "flash" models for my needs. They are a joy and they cost pennies per answer. Great job.
DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
161–170 of 214 posts
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#162>In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price Please don't do this kind of thing. If a user has validated a workflow on V4 Pro, they might not want to suddenly start testing it in production on V4.1 Flash. Instead, keep V4 Pro around but dep…
Usually I would very much agree with you, but those things are not deterministic so if that's an issue for you you're probably not making the right choices.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#163Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#164Earlier quoted context omitted.
Usually I would very much agree with you, but those things are not deterministic so if that's an issue for you you're probably not making the right choices.
That's a narrow take. Non-deterministic doesn't mean random; workflows can be reasonably validated and consistent to some known degree. I work for an education department that serves a chatbot for students, and model changes go through painstaking content safety reviews. I initially assumed it's just a bunch of bureaucratic paranoia. But every other model upgrade has a measurably different adherence to the existing s…
I'm not being a d**, just saying, the problem you have is something that I have faced EXACTLY, and at least here it's not working until you host in house or remote but on raw hardware. Otherwise it keeps having subtle changes, and you will notice no LLM API providers has guarantees about these.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#165Earlier quoted context omitted.
Usually I would very much agree with you, but those things are not deterministic so if that's an issue for you you're probably not making the right choices.
Crossing the street and Russian roulette both have non-deterministic risks of injury. And yet I would be bothered to find out that that on my way to work, I was playing Russian roulette by surprise.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#166> all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price If I'd carefully tested and optimized prompts against Pro I wouldn't be keen on this particular news. I feel like API model providers should lean towards not swapping out models on their paying customers, no matter how much "better" the new model is meant to be.
I imagine they're doing this due to capacity issues or somesuch. They can always relaunch Pro later, meanwhile a little ricered benchmaxxing of their existing flash model provides a temporary cover story. They certainly aren't silly enough to think this won't impact existing Pro users
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#167Earlier quoted context omitted.
I am legitimately more excited for this release than any frontier models at this point. I don't need a model that can invent new mathematics. I need something that is fast, cheap, and consistent. Give me that and I can build and scale.
LLMs are not consistent
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#168>In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price Please don't do this kind of thing. If a user has validated a workflow on V4 Pro, they might not want to suddenly start testing it in production on V4.1 Flash. Instead, keep V4 Pro around but dep…
In this case, Deepseek organization is under a lot of pressure due to compute constraints. It would be better if they just throw a 404 instead of rerouting though so customers are not surprised by subtle changes in behavior.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#169You'd think it would have been something they did a year ago, but here we are. Still.
Re: DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
#170>In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price Please don't do this kind of thing. If a user has validated a workflow on V4 Pro, they might not want to suddenly start testing it in production on V4.1 Flash. Instead, keep V4 Pro around but dep…
If I were paying anthropic prices, I'd expect it, but Deepseek is a super scrappy upstart in comparison and intentionally arbitraging on price. I would never expect them to do that.