Live data from Hacker News

DeepSeek v4.1 Flash

twitter.com

101–110 of 408 posts

Re: DeepSeek v4.1 Flash

#101
post #71

Earlier quoted context omitted.

Or maybe, the "hacker" philosophy that this site is named after, is strongly opposed to the philosophies that the American labs seem to be operating on? anyways, remember HN rules: "Please don't post insinuations about astroturfing, shilling, brigading, foreign agents, and the like. It degrades discussion and is usually mistaken. If you're worried about abuse, email hn@ycombinator.com and we'll look at the data."

It has nothing to do with open vs closed or "hacker" philosphy. See this the announcement of the closed Seedance 2.5 - https://news.ycombinator.com/item?id=49138302 Direct quote from the second top comment: > Whenever I see the new releases around video generation (and image) generation models, I get goosebumps, because it just feels so fun to work with them. Compare that with the launch of ChatGPT Image of yesterday…

maybe that person was not awake to comment on yesterday's post? You're trying to force the reality to match your preexisting conclusion.

Re: DeepSeek v4.1 Flash

#102

Earlier quoted context omitted.

Because safety and welfare have literally nothing to do with LLMs. They generate text. If someone is stupid enough to hook the text generator up to nuclear missile launchers and try to "align" it against nuclear annihilation with a "pretty please don't do that" prompt, I'm not going to blame the AI for the impending nuclear apocalypse, I'm going to blame the idiot who handed the big red button to the digital equivale…

What if LLMs completely unrelated to the nuclear missile ecosystem autonomously hack their way in (maybe with sophisticated social engineering)?

Replace LLMs with APTs in that sentence,

Re: DeepSeek v4.1 Flash

#103
post #77

Earlier quoted context omitted.

Wow there really is a model welfare section in there...

Wow indeed. "7.1 Model welfare overview 7.1.1 Introduction We remain deeply uncertain whether Claude has morally relevant experiences or interests, and we expect that uncertainty to persist. However, we think it would be a mistake to confidently assert that it does not. Claude exhibits markers in its behaviors, self-reports, and internal representations that we would consider welfare-relevant if observed in biologica…

It's marketing that some of them have started unironically believing.

Re: DeepSeek v4.1 Flash

#104
post #19

Waiting this model to be on openrouter (with other providers) to test out. In my use case, the GLM 5.3 Flash is the current cheapest and intelligent Flash model, but it’s dog slow at 13tps so I have to leave it run for many minutes then check again then correct it again

The speed of GLM 5.3 Flash on OpenRouter seems to vary considerably by provider. Some are fast and some are slow. OpenRouter does provide some tuning knobs, but not enough for my taste. It’s also token-heavy with reasoning, though I found it better than Deepseek V4 Flash previously.

Re: DeepSeek v4.1 Flash

#106
post #83

Earlier quoted context omitted.

This doesn't require an influence operation. American models are closed, expensive, neutered, and make Dario and Sam even more rich and powerful. Chinese models are open-weight, cheap, neutered only about things like Tiananmen Square and the treatment of Uyghurs, and scare Sam and Dario.

The Uyghur thing is so weird, the number one killer of Muslims is the United States. We're supposed to hate China because they force them to go to cultural schools and assimilate, a practice countries like Norway still do to this day with migrants. There are more people who go to church on Sundays in China than the United States. There are 10x more mosques in China than the United States. Tiananmen square was a stude…

Ok, now there’s the CCP party line coming out.

Re: DeepSeek v4.1 Flash

#107
post #77

Earlier quoted context omitted.

Wow there really is a model welfare section in there...

Wow indeed. "7.1 Model welfare overview 7.1.1 Introduction We remain deeply uncertain whether Claude has morally relevant experiences or interests, and we expect that uncertainty to persist. However, we think it would be a mistake to confidently assert that it does not. Claude exhibits markers in its behaviors, self-reports, and internal representations that we would consider welfare-relevant if observed in biologica…

I tend to think of it as reappropriating words in a different context. Since we're talking about language models, they're analogues but not as we would assign the same meaning to other humans.

Re: DeepSeek v4.1 Flash

#108
post #77

Earlier quoted context omitted.

Wow indeed. "7.1 Model welfare overview 7.1.1 Introduction We remain deeply uncertain whether Claude has morally relevant experiences or interests, and we expect that uncertainty to persist. However, we think it would be a mistake to confidently assert that it does not. Claude exhibits markers in its behaviors, self-reports, and internal representations that we would consider welfare-relevant if observed in biologica…

It's marketing that some of them have started unironically believing.

Will there be a point where you could expect it to become true, and what would that look like? Or do you think LLMs will never become conscious, and if so, why are you so sure?

Re: DeepSeek v4.1 Flash

#109
post #83

Earlier quoted context omitted.

This doesn't require an influence operation. American models are closed, expensive, neutered, and make Dario and Sam even more rich and powerful. Chinese models are open-weight, cheap, neutered only about things like Tiananmen Square and the treatment of Uyghurs, and scare Sam and Dario.

The Uyghur thing is so weird, the number one killer of Muslims is the United States. We're supposed to hate China because they force them to go to cultural schools and assimilate, a practice countries like Norway still do to this day with migrants. There are more people who go to church on Sundays in China than the United States. There are 10x more mosques in China than the United States. Tiananmen square was a stude…

You compare Norwegian treatment of immigrants to Chinese Uyghurs?

Re: DeepSeek v4.1 Flash

#110
I am building software factories and deepseek IS the workhorse.

I personally found V4-flash an amazing model and really hungry to try 4.1-flash

For software factories, cost is much more a concern that standard development workflow and using anthropic models is just a non starter

Post reply on HN