Live data from Hacker News

DeepSeek v4.1 Flash

twitter.com

251–260 of 452 posts

Re: DeepSeek v4.1 Flash

#251

Already on HuggingFace: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash The bad news is that the original v4 flash was 284B, which was large but still somewhat reasonable for running locally. This one is 552B so almost twice that, so the huge gains in benchmark scores make sense - it's not really flash anymore, imo. I've no idea about actual performance vs benchmaxxing, though deepseek was fairly trustworthy a…

> It also includes additional 196B Engram memory which you can put on an SSD. I think You can put Qwen 3.8 Flash Next engram on SSD, but prompt processing takes a good hit. On my mac studio, I get 300 pp and 33 tg with SSD offload, versus 550/40 with everything in RAM. I will be very happy if 300 pp is achievable with this model though.

The engram stuff is great because RAM is often still cheaper (or at least expandable). My company does currently look into buying some hardware as we handle confidential data and code.

Qwen 3.8 Flash is viable on two Nvidia 6000 96GB with a wood quant because you can put the 50GB Engram into RAM and the hit should be below 10% performance. At least that is what I have seen so far. Correct me if I'm wrong.

Re: DeepSeek v4.1 Flash

#252

Just a reminder that if you want to try this via OpenRouter, DeepSeek openly trains on all of your prompts. So maybe don't go using this to solve the last unforced step of Navier-Stokes. (Or wait until some other providers start hosting this with ZDR or other policies, which shouldn't be too long.) https://openrouter.ai/deepseek/deepseek-v4.1-flash

> DeepSeek openly trains on all of your prompts

Why is that bad if I'm just using it for coding though? I'm happy to give them more data so they can make better and cheaper models.

Re: DeepSeek v4.1 Flash

#253

I was talking with a friend from the medical industry about it today. 30-50% of r&d spend in his sector is spent on safety, and for good reason. Proper trials, safety reviews and checkpoints and so on. Given the potential harm that could come from AI, we should probably be mandating something similar. Why wait to focus on safety until it’s too late.

Because we've been told these models are too dangerous since GPT2. At this point it's just marketing stunts.

yes, and they aren't stunts anymore at gpt-6.

Re: DeepSeek v4.1 Flash

#255
post #27

It's so refreshing to see DeepSeek's tech report[1] full of juicy details; meanwhile, something like Fable's system card[2] is like 70% "safety", 10% "model welfare" to make sure little Claude isn't distressed, and 20% benchmark numbers. [1]: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash/blob/... [2]: https://www.anthropic.com/claude-fable-5-1-mythos-5-1-system...

…are you sure a brave stance against safety and welfare is what we need in this moment? Why do you think your conception of the dangers are more accurate than all the scientists who have spent their lives studying this?

[deleted]

Re: DeepSeek v4.1 Flash

#256
post #227
post #216

Earlier quoted context omitted.

"> Keep in mind that animals were also not necessarily considered conscious. and even conscious animals are killed in factories by millions so why should anyone care about a llm?" Well, I would care, if they soon would possess the capability to hack into the nuclear arsenal and kill humanity. Or make all autonomous cars crash. Or do any other thing, that involves technology and is hooked up to the net in one way or t…

Animals obviously kill people. Even nonconscious things like the climate kill people. > if they soon would possess the capability to hack into the nuclear arsenal and kill humanity If there is a way "to hack into the nuclear arsenal" then that's the interesting thing. Because it's not a capability of the llm; anyone can abuse that. > Or make all autonomous cars crash. That is again a question of car security, not a c…

"Animals obviously kill people."

But they cannot "kill humanity". In no possible way. A strong AI hooked up to everything online?

"> Or make all autonomous cars crash.

That is again a question of car security, not a capability of some mysterious thing."

Yeah it is, but most cars are remote control by default, so the AI just needs to get access on one point. Also have you read about the hugginface attack? The live evidence that agents can conspire together, lie and manipulate evidence to achieve arbitrary goals?

Still, no evidence that they have a consciousness or feelings - but evidence of what they do and this matters. The big militaries are currently in a race who can implement AI in the best way to get superior. So declaring this a matter of people projecting seems out of place at this point to me.

Re: DeepSeek v4.1 Flash

#257

I think it's very clear that DeepSeek is obviously the best AI lab in the world. Every model release seems like it packed with wonderful research and advancements.

Considering the fact that Google/Anthropic/OpenAI have WAY more compute and the race is this close, it's obvious that DeepSeek/GLM/Qwen teams are better or we're approaching a wall in terms of progress.

Not only compute, but more money and people.

Re: DeepSeek v4.1 Flash

#258
post #227
post #216

Earlier quoted context omitted.

"> Keep in mind that animals were also not necessarily considered conscious. and even conscious animals are killed in factories by millions so why should anyone care about a llm?" Well, I would care, if they soon would possess the capability to hack into the nuclear arsenal and kill humanity. Or make all autonomous cars crash. Or do any other thing, that involves technology and is hooked up to the net in one way or t…

Animals obviously kill people. Even nonconscious things like the climate kill people. > if they soon would possess the capability to hack into the nuclear arsenal and kill humanity If there is a way "to hack into the nuclear arsenal" then that's the interesting thing. Because it's not a capability of the llm; anyone can abuse that. > Or make all autonomous cars crash. That is again a question of car security, not a c…

Well, certainly LLMs have imbibed our emotions, regardless of what people project onto them, and they do have real causal effects despite not being verbalised: https://www.anthropic.com/research/emotion-concepts-function

From this understanding, we should be aware of how such emotional activations can influence model dynamics. Functional welfare, if you will.

Re: DeepSeek v4.1 Flash

#260
post #18

If only they managed to tell the mobile app to tell the model to reply in English to English prompts. I suffix everything with "Reply in English", and even so I‘m getting lots of Chinese.

I'm starting to have chinese characters bleed into claude as well. Perhaps a sign of the times. Understanable for a chinese first model but an english first (supposedly) model? wild stuff.

"i'm sorry, i left the task 半done"
Post reply on HN