Sounds like a big improvement. No mention of weights, just API. When will the updated weights be released?
DeepSeek-V4-Flash Update
261–270 of 362 posts
Re: DeepSeek-V4-Flash Update
#262Why didn't they increment the version as atleast a patch update
Re: DeepSeek-V4-Flash Update
#263Deepseek and moonshot are the only two providers I consent to training for.
CCP will be happy! Go on and share all your data with them..
US or American, don't trust anyone, these are open weight models. Host them yourself if you feel strongly about privacy (as you should). I honestly don't know of any OSS open weight models from the US labs as good as Kimi K3 or Deepseek V4 though.
Re: DeepSeek-V4-Flash Update
#264Earlier quoted context omitted.
Sounds like it'll replace v4-flash, v4.1 would be nice to keep both available. On the other hand, it's nice to just get an improvement on anything that asks for "deepseek-v4-flash" without having to change the model string.
I think that's backwards. Anything that changes the performance of a model deserves a minor version bump. A new model has to be qualified before being pushed to production; but we don't get the choice here, just cross your fingers there are no regressions at all on all possible tasks the model might be asked to do.
Re: DeepSeek-V4-Flash Update
#265What was preventing them from calling it v4.1-Flash to distinguish it better?
Re: DeepSeek-V4-Flash Update
#266Earlier quoted context omitted.
Recommendation? No. Just go with the passive-aggressive advice "let pi build it for you". :-) To be more constructive, what I did (as an experiencd SWE but a complete noob to agentic coding): went to pi.dev's extension marketplace and looked into all the new shiny stuff. Subagents, mcps, context and memory optimizers, skills. Using the most popular ones (not necessarily the best ones) It was like 15years ago learning…
I found that opencode absolutely lets you do everything that pi does. With a few niceties in the UI on top.
Re: DeepSeek-V4-Flash Update
#267This is more exciting than k3, IMO. Dsv4 models are extremely cheap to serve. Improving their capabilities has lots of downstream effects, as it becomes "good enough" for more and more tasks. DS was serving the pro version at extremely low prices for a long time, and they've had integrations with opencode & other providers, so they likely gathered a lot of data from real developers doing real tasks (on openrouter the…
Starting to wonder if the free big pickle model on opencode has been DSV4F0731 for the past few months. It’s been incredibly fast and good.
Re: DeepSeek-V4-Flash Update
#268I use deepseek for a lot of my personal day-to-day agent needs, and I will simply put this here and let this speak for itself, last 30 days: - Cost: $4.55USD - API requests: 3,467 - Tokens: 323,183,886 And as an engineer who leads a small team, I have very high standards for quality, and these carry across to my personal projects where I use deepseek. It has not disappointed at all for coding or review tasks. For eve…
Re: DeepSeek-V4-Flash Update
#269I use deepseek for a lot of my personal day-to-day agent needs, and I will simply put this here and let this speak for itself, last 30 days: - Cost: $4.55USD - API requests: 3,467 - Tokens: 323,183,886 And as an engineer who leads a small team, I have very high standards for quality, and these carry across to my personal projects where I use deepseek. It has not disappointed at all for coding or review tasks. For eve…
Re: DeepSeek-V4-Flash Update
#270Earlier quoted context omitted.
Starting to wonder if the free big pickle model on opencode has been DSV4F0731 for the past few months. It’s been incredibly fast and good.
At least in the past it was GLM-4.6. IDK if it is ever changed. https://github.com/anomalyco/opencode/issues/4276