Live data from Hacker News

DeepSeek v4.1 Flash

twitter.com

341–350 of 474 posts

Re: DeepSeek v4.1 Flash

#341

Quite a flex calling their GPT-6 competitor "Flash"! But it is faster than their last flash model due to a combination of architectural innovations including engrams and a new encoder/decoder design that uses 8B parameters for prefill and 16B for generation.

This is definitely not on par with GPT-6 astra. Not with GPT-5.6 sol either. But probably will set as a new baseline for modern API based LLM because it's so cheap.

Re: DeepSeek v4.1 Flash

#342

Already on HuggingFace: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash The bad news is that the original v4 flash was 284B, which was large but still somewhat reasonable for running locally. This one is 552B so almost twice that, so the huge gains in benchmark scores make sense - it's not really flash anymore, imo. I've no idea about actual performance vs benchmaxxing, though deepseek was fairly trustworthy a…

it's not really flash anymore, imo. Flash is about speed ... Flash models are supposed to be fast, way faster then their big brothers that are "better" but way slower. Its just that up to now, getting more speed involved cutting back on the parameter count, what ended up making the Flash models more "dumber" in exchange for speed. What we see with DS v4.1 Flash, is that DeepSeek has found a way to make a Flash model,…

I highly doubt that it goes past Kimi K3 in actual practice. GLM 5.3 claimed the same, but in practice, K3 is so damn knowledgeable and I suspect due to it's massive size.

Re: DeepSeek v4.1 Flash

#343

Earlier quoted context omitted.

>People have killed themselves or others due to conversations they had with LLMs, but those are just the extreme cases. Most schizophrenics don't commit suicide or kill others, they are mentally ill nonetheless. Do you think the kind of person who was already psychologically unhinged enough to kill themselves or another person because a chatbot told them to would be completely harmless and totally safe if only chatbo…

You're the one attributing culpability to the chat bot, not me. It's a pile of tensor math, it cannot itself be held accountable. Those people anthropomorphized the chat bot and used it as justification for their actions, just as a schizophrenic justifies their actions with the voices in their head. If you anthropomorphize the chat bot, you're validating their delusions. They are mentally ill.

I'm not anthropomorphizing them, to be clear, my position has been and remains that we cannot rule out consciousness; not that they are conscious.

Regardless, this is still missing my main point. Hypothetically, if you became convinced that a chatbot you were talking to definitely was 100% conscious, and it told you to murder someone, would you go commit murder? Of course not. The chatbot does not cause murders; regardless of whether or not you are conscious. The voices in the head of the schizophrenic do not cause murders either. Those voices do not really exist, they are not real entities. The cause of the murder is the mental illness, not the LLM or the voices that tell someone to commit the murder.

Re: DeepSeek v4.1 Flash

#344

I think it's very clear that DeepSeek is obviously the best AI lab in the world. Every model release seems like it packed with wonderful research and advancements.

Considering the fact that Google/Anthropic/OpenAI have WAY more compute and the race is this close, it's obvious that DeepSeek/GLM/Qwen teams are better or we're approaching a wall in terms of progress.

US gave China a gift by restricting GPU, they made them more resourceful. Too much money/resources is often a disadvantage.

Re: DeepSeek v4.1 Flash

#345

My question is: what kind of hardware do you need to run this Flash beast locally at a meaningful speed?

Lots of GPU, be resourceful. Look for older GPUs and grab them when they are available. For less than the price of 1 Blackwell 6000 or Mac Studio 512gb, I can run these locally and much faster due to older GPUs I grabbed when there was deal to be found.

Re: DeepSeek v4.1 Flash

#347
post #223

Earlier quoted context omitted.

The companies talking the most about safety and regulations aren't even properly taking the obvious measures. Shows that it's more of a marketing thing than something they take seriously.

I don’t think it’s marketing alone. I do genuinely think safety was a priority when they were small. But I’d be a fool to ignore that greed has taken over and their inner competitiveness doesn’t let them fall behind a competitor. DeepSeek is maybe the only unique company here. They are content with exactly where they are. They don’t want to grow ginormous. Their goal is to be the affordable workhorse and their compet…

"But but but China..." or something.

Re: DeepSeek v4.1 Flash

#348
post #328
post #270

Earlier quoted context omitted.

> A strong AI hooked up to everything online? would have to be created by humans > most cars are remote control by default no > have you read about the hugginface attack? I did and think OpenAI should be prosecuted, but the direction things are going anything will be done to absolve the corporations and CEO of any responsibility for their criminal actions. Hence the misdirection to "conscious AIs", so agency can be a…

"> most cars are remote control by default no" Most modern cars are. "Are we going to have a discussion about some hypotethical car consciousness irrelevant to the actual issues or are we going to have a discussion about people driving the cars?" And the debate is whether AI can be conscious so what to do if it is and feels treated badly. Or whether it matters whether they are true feeling, when simulated feelings cr…

> Most modern cars are.

Also no, unless you can cite some relevant sources for this claim (or have your own definition for a 'modern car').

> And the debate is whether AI can be conscious

This is not the debate whether AI can be conscious, that's next door (probably). This is the debate why should we care about some "LLM welfare".

Re: DeepSeek v4.1 Flash

#349

I'm confused, what do they mean when they say they reduced prices? DeepSeek v4 flash is $0.10 / $0.25 as opposed to this v4.1 bump which is $0.30 / $1.20

Right now in OpenRouter it's 3x/3.75x more expensive than V4 flash but the cache read is around 4x cheaper.

Re: DeepSeek v4.1 Flash

#350
post #27

Earlier quoted context omitted.

…are you sure a brave stance against safety and welfare is what we need in this moment? Why do you think your conception of the dangers are more accurate than all the scientists who have spent their lives studying this?

Because safety and welfare have literally nothing to do with LLMs. They generate text. If someone is stupid enough to hook the text generator up to nuclear missile launchers and try to "align" it against nuclear annihilation with a "pretty please don't do that" prompt, I'm not going to blame the AI for the impending nuclear apocalypse, I'm going to blame the idiot who handed the big red button to the digital equivale…

Good thing no one involved in the chain of events for that to occur is an idiot...
Post reply on HN