Live data from Hacker News

DeepSeek v4.1 Flash

twitter.com

461–470 of 541 posts

Re: DeepSeek v4.1 Flash

#461

Earlier quoted context omitted.

What's your definition of sentient? Or, maybe more precisely, consciousness? I think it's reasonable to at least start thinking about these questions. It has long been established that LLMs have good theory of mind [1]. And there is a bunch of empirical research about all sorts of capabilities that we typically associate with consciousness [2], like identity [3] and metacognition [4]. The METR report shows agents sac…

Not the same person but to me, the answer is that it does not matter, and that all these attempts at making it matter are pure marketing and emotional manipulation. It's not a living creature. It's an autoregressive pure function of token-sequence to token, which is capable of incredible things, but it's still just a function. It is not alive as it cannot die in any meaningful sense. It is less "alive" than the RNA m…

Does being made out of meat instead of silicon make you more sentient?

Re: DeepSeek v4.1 Flash

#462
post #214

Earlier quoted context omitted.

While typical investors in their last round are subject to a five-year lock-up and will not have voting rights, China's National Artificial Intelligence Industry Investment Fund also put money into it, retaining both voting rights and freedom from the lock-up. Nothing really new if you're aware of how involved the CCP is with companies of strategic importance in China. https://www.reuters.com/world/asia-pacific/china…

Oh of course, you’re not getting into positions of power by not playing by the party’s rules. And if you get notions that you can tell THEM what to do you’ll be swiftly dealt with. The company is doing well and providing great PR so the party is content to not meddle too much I imagine. My comparison with American labs is more that I think they have to deal with bean counters, creditors, investors etc which can shuff…

> My comparison with American labs is more that I think they have to deal with bean counters, creditors, investors etc which can shuffle incentives and aims (and is a big reason why they dont do open weights anymore)

And don't forget kowtowing to Trump.

Re: DeepSeek v4.1 Flash

#463
post #207

Earlier quoted context omitted.

> welfare We should be paying attention to it just in case it ends up mattering enormously. It's cheap insurance. Aside from that, US labs' system cards have been pretty useless for a while—I think the last great one was the combined system card for Claude 4 Sonnet and Opus.

> We should be paying attention to it just in case it ends up mattering enormously. It's cheap insurance. This sounds a lot like the argument some people give for praying and going to church even if you aren't a believer. "You should be doing it just in case God ends up being real."

On the other hand, we have recently taught rocks to think about software engineering. And while not perfect, they're surprisingly good at it. Once you start building things that are even a little bit like minds, I suspect that it's worthwhile to consider that the future might end up looking a bit like science fiction.

The alternative is to insist that Nothing Ever Happens, and the future won't get too weird. Which is no longer a bet I'm entirely comfortable with. Weirdness is at least a possibility.

Re: DeepSeek v4.1 Flash

#465
post #262

Earlier quoted context omitted.

I believe consciousness is necessarily stateful. The LLM itself (ignoring implementation details that don't change the results) is a deterministic pure function. It's functionally equivalent to an enormous lookup table. If I accepted LLMs as conscious, then I would have to accept panpsychism, which I do not, and which most other humans also act as though they do not.

I don't think so because any stateful function can be made stateless just by making its state an input, and vice versa. They're mathematically equivalent, so it would be super weird if it had any implications for consciousness.

This assumes the functionality of brains can be fully captured as a deterministic mathematical function, but the function of the brain may well depend on nondeterministic quantum states that can't be reduced to stateless functions: https://en.wikipedia.org/wiki/Quantum_mind

Re: DeepSeek v4.1 Flash

#466
post #376

Earlier quoted context omitted.

>it really amazes me how fearless Deepseek are Reel it in a bit, man.

The circle jerking of Chinese models on this site never ceases to amuse me.

Open is better than closed, simple as. Nothing to do with China versus America, I'll always stan the open models and cast my aspersions on the closed ones.

Re: DeepSeek v4.1 Flash

#467
post #376

Earlier quoted context omitted.

>it really amazes me how fearless Deepseek are Reel it in a bit, man.

The circle jerking of Chinese models on this site never ceases to amuse me.

>The circle jerking of Chinese models on this site never ceases to amuse me.

As opposed to "ask the model about Tiananmen" which seems to be the site's favorite pastime about Chinese models. ;)

--

Sarcasm aside, I don't think people are happy about _Chinese_ models making advances. They are happy about _open_ models making advances. It's just coincidental that China is the one making them.

If some American lab were to develop a SOTA open model most people here will be equally excited. Although besides GPT-OSS-120B the American labs have been disappointing in this regard.

Re: DeepSeek v4.1 Flash

#468
post #398

DeepSeek Harness, install it, thank me later. You won't believe the productivity gains for just pennies https://deepseek.com/harness/en/

How does it compare against the Pi harness, which I thought was the unofficial harness champion so far, in your workloads?

Re: DeepSeek v4.1 Flash

#469
post #274

Earlier quoted context omitted.

No language models are programmed, they are "grown" or evolved from data. There's no print statements or human entered logic involved in the raw model expression at all. The only thing that humans have programmed is efficient parallel dot product pipelines that "animate" (for lack of a better word) the models. Everything these models do is emergent from their backpropgation guided evolution. This even includes in con…

You have completely misunderstood what I was saying so badly I can't even formulate a response other than to suggest you read my reply again. I was not suggesting that LLMs are programmed with print statements, for fuck's sake.

This perspective that consciousness cannot be programmed can only make sense if you're a dualist. We don't know how consciousness arises. If you're a naturalist it can't be ruled out based on the simplicity of the algorithm.

Re: DeepSeek v4.1 Flash

#470
post #27

Earlier quoted context omitted.

…are you sure a brave stance against safety and welfare is what we need in this moment? Why do you think your conception of the dangers are more accurate than all the scientists who have spent their lives studying this?

Because safety and welfare have literally nothing to do with LLMs. They generate text. If someone is stupid enough to hook the text generator up to nuclear missile launchers and try to "align" it against nuclear annihilation with a "pretty please don't do that" prompt, I'm not going to blame the AI for the impending nuclear apocalypse, I'm going to blame the idiot who handed the big red button to the digital equivale…

So if a model with exactly the same architecture controls a robot then it's suddenly sentient or what? "They generate actions in the real world"
Post reply on HN