Live data from Hacker News

DeepSeek-V4-Flash Update

api-docs.deepseek.com

311–320 of 362 posts

Re: DeepSeek-V4-Flash Update

#311
Long story, I have humongous zai GLM 5.2 token budget that I'm using in a similar fashion as many comments explain here. GPT 5.6 or Fable 5 for planning, GLM for implementing, researching, extracting and many other tasks I consider grunt work. Very happy with the performance, speed isn't all that good though. I'd be curious to hear from someone who works with both, DeepSeek and GLM side by side.

Re: DeepSeek-V4-Flash Update

#312
post #167

I wonder when the antirez/ds4 group will have an update to their high accuracy 2 bit quant. Although it's funny that I am thinking about that at all because I have a 2060 :P . My local inference is playing with Gemma 4 E2B and MiniCPM 5 1B.

Looks like a couple of people are already on it: https://github.com/antirez/ds4/issues/635

Re: DeepSeek-V4-Flash Update

#313

Earlier quoted context omitted.

He explicitly said he wants open weight models at his recent speech at an AI conference in Shanghai: https://news.ycombinator.com/item?id=48970449#48970784

The Chinese equivalent of the expression 'open weights' appears nowhere in this talk. You believed the Western press. Chinese AIs are not typically open. The most important by far is Bytedance.

The second China shuts down open weights is the second they lose the AI race. Nobody is going to use Chinese models if they're closed, besides hobbyists who don't have anything they care about getting stolen.

Re: DeepSeek-V4-Flash Update

#315
post #23

Oh my goodness what an update. I need these weights. It's an incredible model for the size. The improved tool calling etc. should be able to make my harness way simpler. This runs at mega-speed on prosumer hardware (2x RTX Pro 6000).

How many tok/s are we talking about?

Re: DeepSeek-V4-Flash Update

#316
post #287

Can someone please explain how these models aren’t just fine tuned for benchmarks? I’m not plugged in to this space much but it seems like such an obvious problem…

They definitely are - but also people are using them pretty extensively for work. So ultimately you can't really fake "is it good". But there's no real measurements of that when a model is released, so we are stuck with benchmarks.

I will continue to ignore the benchmarks.

Re: DeepSeek-V4-Flash Update

#317
post #201

Earlier quoted context omitted.

Sounds like it'll replace v4-flash, v4.1 would be nice to keep both available. On the other hand, it's nice to just get an improvement on anything that asks for "deepseek-v4-flash" without having to change the model string.

I think that's backwards. Anything that changes the performance of a model deserves a minor version bump. A new model has to be qualified before being pushed to production; but we don't get the choice here, just cross your fingers there are no regressions at all on all possible tasks the model might be asked to do.

its an open source model if it is important enough that you have to worry about a new version breaking something then why the fuck was that not running on your own servers this is not an Anthropic or OpenAI closed model that you only have access through an api

Re: DeepSeek-V4-Flash Update

#318

Earlier quoted context omitted.

Weird, I'm using 5.6 Sol through Chinese resellers and it reverse engineers stuff just fine

I think I’m missing something, what do you mean Chinese resellers? As I understand it, it’s difficult to even access OpenAI in China, how could they be reselling it? Do you mean something like openrouter but a Chinese version or something?

They usually resell codex subscriptions as api so it's cheaper than the official api

Re: DeepSeek-V4-Flash Update

#319

Earlier quoted context omitted.

Weird, I'm using 5.6 Sol through Chinese resellers and it reverse engineers stuff just fine

Do you have any suggestions for resellers? My Gmail username is the same as my HN username if you're not comfortable posting that here. Thanks.

Here's a big list: https://vectoral.com/blog/token-relay-market

Re: DeepSeek-V4-Flash Update

#320

Earlier quoted context omitted.

Do you have any suggestions for resellers? My Gmail username is the same as my HN username if you're not comfortable posting that here. Thanks.

Here's a big list: https://vectoral.com/blog/token-relay-market

Thank you
Post reply on HN