Live data from Hacker News

DeepSeek v4.1 Flash

twitter.com

111–120 of 464 posts

Re: DeepSeek v4.1 Flash

#111
post #83

Earlier quoted context omitted.

This doesn't require an influence operation. American models are closed, expensive, neutered, and make Dario and Sam even more rich and powerful. Chinese models are open-weight, cheap, neutered only about things like Tiananmen Square and the treatment of Uyghurs, and scare Sam and Dario.

The Uyghur thing is so weird, the number one killer of Muslims is the United States. We're supposed to hate China because they force them to go to cultural schools and assimilate, a practice countries like Norway still do to this day with migrants. There are more people who go to church on Sundays in China than the United States. There are 10x more mosques in China than the United States. Tiananmen square was a stude…

I simply mean the Chinese models will refuse sensitive domestic issues, which are unlikely to affect the average user's work, while American models refuse things that can limit their utility, e.g. how the HF team had to investigate the openai attack with Chinese models because the American models refused.

(I mainly mentioned those specific topics to establish clearly I am not part of the alleged influence operation.)

Re: DeepSeek v4.1 Flash

#112
post #16

I'm a big fan of DeepSeek. Also, ask it what model it is :) In Pi (pi.dev), it tells me it's definitely Claude by Anthropic, via the API via curl it tells me it's "probably ChatGPT", its very funny.

Definitely not Claude, deepseek is too fast, so I bet it's ChatGPT. :P

Re: DeepSeek v4.1 Flash

#113
post #27

Earlier quoted context omitted.

…are you sure a brave stance against safety and welfare is what we need in this moment? Why do you think your conception of the dangers are more accurate than all the scientists who have spent their lives studying this?

> scientists who have spent their lives studying this Please point me to one actual accredited scientist who has spent a lifetime studying AI alignment? Pretty much this whole field is only 5 years old

The field is much older, MIRI is ~20 years old. Look up Eliezer Yudkowsky.

Re: DeepSeek v4.1 Flash

#114
OpenCode Go is currently running a 4x usage promo on DeepSeek v4.1 flash, not a bad way to get your feet wet (even if their cache hit prices are probably still very sub-optimal)

Re: DeepSeek v4.1 Flash

#115

Earlier quoted context omitted.

Excuse me for not being interested in over 100 pages of how well the model can refuse and block my requests, especially considering how fun it is to waste my time trying to get around those restrictions when they inevitably trigger because the clanker thinks that I'm doing something naughty, all the while it can't reliably center the proverbial div without doing something stupid itself.

Meanwhile I have an uncensored qwen 3.8 27B here that will happily attempt to (as a crude and randomly chosen sampling of bad/evil things) give me the recipes for meth, how to make an IED, write a manifesto in support of a horrible ideology, or commit various forms of fraud. Now I certainly wouldn't recommend that anyone try to follow what it says to do, because it's almost certainly very wrong on key parts that woul…

Yep. Just like a kitchen knife will make no attempt to prevent me from stabbing anyone with it.

Here's a dirty secret though -- you don't actually need an abliterated/uncensored version of the model to get it to do this. I can do this with every and each open weight model, as served from OpenRouter, using vanilla model weights.

Re: DeepSeek v4.1 Flash

#116

Earlier quoted context omitted.

> scientists who have spent their lives studying this Please point me to one actual accredited scientist who has spent a lifetime studying AI alignment? Pretty much this whole field is only 5 years old

The field is much older, MIRI is ~20 years old. Look up Eliezer Yudkowsky.

The field was purely theoretical 20 years ago, and Yudkowsky is pretty much the dictionary definition of "not accredited"

Re: DeepSeek v4.1 Flash

#117

I think it's very clear that DeepSeek is obviously the best AI lab in the world. Every model release seems like it packed with wonderful research and advancements.

[flagged]

> Yes comrade, they are the best.

> Did you do your daily data centers errrr baaaaddd AI generated post for Facebook?

Please stop insulting people. I'm all for heated discussion, but you are not discussing, you insult.

Now go away, before your insults come back to you, "comrade from Facebook".

Re: DeepSeek v4.1 Flash

#118

First flash model with multimodal support? I think Flash series might be the main focus going forward for them. Tried it out and it’s better than v4 pro

No. DSv4-Flash-Vision-Exp is what I use and it has vision.

Re: DeepSeek v4.1 Flash

#119

Earlier quoted context omitted.

Wow there really is a model welfare section in there...

To me it reads like pure propaganda. Anthropic really wants us to think that they've made something sentient. I think that's really dangerous.

Marketing, like Volvo cars being safer etc

Re: DeepSeek v4.1 Flash

#120
This seems like a very nice release. Just ran it over my Kubernetes security benchmark that I run for most new releases. It was fast, cheap, and got a high scoring result, nice!
Post reply on HN