Live data from Hacker News

DeepSeek-V4-Flash Update

api-docs.deepseek.com

261–270 of 362 posts

Re: DeepSeek-V4-Flash Update

#263
post #232
post #91

Deepseek and moonshot are the only two providers I consent to training for.

CCP will be happy! Go on and share all your data with them..

Yeah, because, your data is best collected by the US labs, right?

US or American, don't trust anyone, these are open weight models. Host them yourself if you feel strongly about privacy (as you should). I honestly don't know of any OSS open weight models from the US labs as good as Kimi K3 or Deepseek V4 though.

Re: DeepSeek-V4-Flash Update

#264
post #201

Earlier quoted context omitted.

Sounds like it'll replace v4-flash, v4.1 would be nice to keep both available. On the other hand, it's nice to just get an improvement on anything that asks for "deepseek-v4-flash" without having to change the model string.

I think that's backwards. Anything that changes the performance of a model deserves a minor version bump. A new model has to be qualified before being pushed to production; but we don't get the choice here, just cross your fingers there are no regressions at all on all possible tasks the model might be asked to do.

That's on you for using a preview/beta. It was properly qualified when it was released. No one complains when ios goes from beta to release; even though it will be ios 27 beta -> ios 27 next month

Re: DeepSeek-V4-Flash Update

#266
post #65

Earlier quoted context omitted.

Recommendation? No. Just go with the passive-aggressive advice "let pi build it for you". :-) To be more constructive, what I did (as an experiencd SWE but a complete noob to agentic coding): went to pi.dev's extension marketplace and looked into all the new shiny stuff. Subagents, mcps, context and memory optimizers, skills. Using the most popular ones (not necessarily the best ones) It was like 15years ago learning…

I found that opencode absolutely lets you do everything that pi does. With a few niceties in the UI on top.

Opencode doesn't even allow providing your own system prompt.

Re: DeepSeek-V4-Flash Update

#267

This is more exciting than k3, IMO. Dsv4 models are extremely cheap to serve. Improving their capabilities has lots of downstream effects, as it becomes "good enough" for more and more tasks. DS was serving the pro version at extremely low prices for a long time, and they've had integrations with opencode & other providers, so they likely gathered a lot of data from real developers doing real tasks (on openrouter the…

Starting to wonder if the free big pickle model on opencode has been DSV4F0731 for the past few months. It’s been incredibly fast and good.

At least in the past it was GLM-4.6. IDK if it is ever changed.

https://github.com/anomalyco/opencode/issues/4276

Re: DeepSeek-V4-Flash Update

#268
post #50

I use deepseek for a lot of my personal day-to-day agent needs, and I will simply put this here and let this speak for itself, last 30 days: - Cost: $4.55USD - API requests: 3,467 - Tokens: 323,183,886 And as an engineer who leads a small team, I have very high standards for quality, and these carry across to my personal projects where I use deepseek. It has not disappointed at all for coding or review tasks. For eve…

OpenRouter?

Re: DeepSeek-V4-Flash Update

#269
post #50

I use deepseek for a lot of my personal day-to-day agent needs, and I will simply put this here and let this speak for itself, last 30 days: - Cost: $4.55USD - API requests: 3,467 - Tokens: 323,183,886 And as an engineer who leads a small team, I have very high standards for quality, and these carry across to my personal projects where I use deepseek. It has not disappointed at all for coding or review tasks. For eve…

You are using DeepSeek's services directly? Doesn't that end up sending at least snippets/chunks of code to a server where it is subject to Chinese government data access laws? Even if I was okay with that, my organization would never be. And even if they were, our partners/vendors/customers would not be. I think that's the sticking point for a lot of people.

Re: DeepSeek-V4-Flash Update

#270

Earlier quoted context omitted.

Starting to wonder if the free big pickle model on opencode has been DSV4F0731 for the past few months. It’s been incredibly fast and good.

At least in the past it was GLM-4.6. IDK if it is ever changed. https://github.com/anomalyco/opencode/issues/4276

I’ve heard that too, but since it’s a stealth model, I suspect they change it to whatever preview model provider that offers a free endpoint. In the past few months, I’ve gotten 4xx errors identifying the provider as DeepSeek.
Post reply on HN