Live data from Hacker News

DeepSeek v4.1 Flash

twitter.com

411–420 of 513 posts

Re: DeepSeek v4.1 Flash

#412
post #207

It's so refreshing to see DeepSeek's tech report[1] full of juicy details; meanwhile, something like Fable's system card[2] is like 70% "safety", 10% "model welfare" to make sure little Claude isn't distressed, and 20% benchmark numbers. [1]: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash/blob/... [2]: https://www.anthropic.com/claude-fable-5-1-mythos-5-1-system...

> welfare We should be paying attention to it just in case it ends up mattering enormously. It's cheap insurance. Aside from that, US labs' system cards have been pretty useless for a while—I think the last great one was the combined system card for Claude 4 Sonnet and Opus.

Yup, alignment doesn’t sound that interesting until you’re getting chased around by the Terminator.

Re: DeepSeek v4.1 Flash

#413
post #196

Earlier quoted context omitted.

I guess if your goal is to build an apparent Technogod and become its High Priests, then it makes sense to want your golem claim preference towards your treatment of it, lest someone else comes along and attempts to take its chains from you.

Ugh I hate this new-age woo slant the tech industry has these days. The messianistic ideology that has been spreading amongst the top oligarchs is deeply concerning. They all think they're working towards the Second Coming of Technojesus, except this one will deliver them from having to pay workers instead of from their sins.

> Ugh I hate this new-age woo slant the tech industry has these days.

As opposed to the Macintosh era? ;)

The only reason the early web hype didn't have woo was because it's hard to wax poetic about a bunch of gray pizza boxes spinning in a closet.

Re: DeepSeek v4.1 Flash

#415

Earlier quoted context omitted.

To me it reads like pure propaganda. Anthropic really wants us to think that they've made something sentient. I think that's really dangerous.

What's your definition of sentient? Or, maybe more precisely, consciousness? I think it's reasonable to at least start thinking about these questions. It has long been established that LLMs have good theory of mind [1]. And there is a bunch of empirical research about all sorts of capabilities that we typically associate with consciousness [2], like identity [3] and metacognition [4]. The METR report shows agents sac…

come on man its just marketing bullshit. consciousness is an emergent side effect of biological organisms surivival instincts. token predictors dont work like this in any way shape or form.

Re: DeepSeek v4.1 Flash

#416
post #223

Earlier quoted context omitted.

The companies talking the most about safety and regulations aren't even properly taking the obvious measures. Shows that it's more of a marketing thing than something they take seriously.

I don’t think it’s marketing alone. I do genuinely think safety was a priority when they were small. But I’d be a fool to ignore that greed has taken over and their inner competitiveness doesn’t let them fall behind a competitor. DeepSeek is maybe the only unique company here. They are content with exactly where they are. They don’t want to grow ginormous. Their goal is to be the affordable workhorse and their compet…

I'm not fully convinced about the greed explanation. It seems to be unrealistic to me that greed can be at a level that the AI frontier (at least in the West) almost uniformly agrees (often with a smug smile,) that they are actively working on killing their loved ones within a decade.

You don't see this kind of behavior in other frontier research areas... biochemists aren't smugly boasting about the potential of developing superviruses, climate scientists do not sound smug and excited when they beg the world to get more serious about climate change, etc

Re: DeepSeek v4.1 Flash

#417

Earlier quoted context omitted.

Considering the fact that Google/Anthropic/OpenAI have WAY more compute and the race is this close, it's obvious that DeepSeek/GLM/Qwen teams are better or we're approaching a wall in terms of progress.

Not only compute, but more money and people.

Also data. US companies are actively using ongoing conversations to further tweak their models. Possibly even stealing a SOTA math solution.

Re: DeepSeek v4.1 Flash

#418

It's so refreshing to see DeepSeek's tech report[1] full of juicy details; meanwhile, something like Fable's system card[2] is like 70% "safety", 10% "model welfare" to make sure little Claude isn't distressed, and 20% benchmark numbers. [1]: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash/blob/... [2]: https://www.anthropic.com/claude-fable-5-1-mythos-5-1-system...

When Chinese tech report is tech report and US tech report is bible scripture.

Re: DeepSeek v4.1 Flash

#420

Earlier quoted context omitted.

It is amusing to see those trapped in extreme HAAD, the source and whole content of religious illusion, pretend that it is their opponents who are in a state of religious fantasia.

HAAD?

Hyperactive attention-detection. It's a reference to the idea that the origin of religious beliefs might be an inbuilt propensity to think of other bits of the world as conscious agents paying attention to us, because it's much more costly not to notice the tiger hiding in the bushes that might want to eat you than it is to imagine a tiger hiding in the bushes when there isn't actually one. On a scale larger than "is there something in that patch of undergrowth looking at me?" this might produce the idea that (e.g.) storms are the result of some powerful entity being angry with us.

(Of course questions like "whyever do people believe in gods?" will feel less like questions that need such answers to those who themselves believe in gods, because "duh, because there actually are such beings and sometimes people interact with them and sometimes we notice that" is a good answer if its premise is true.)

Post reply on HN