Earlier quoted context omitted.
Or maybe, the "hacker" philosophy that this site is named after, is strongly opposed to the philosophies that the American labs seem to be operating on? anyways, remember HN rules: "Please don't post insinuations about astroturfing, shilling, brigading, foreign agents, and the like. It degrades discussion and is usually mistaken. If you're worried about abuse, email hn@ycombinator.com and we'll look at the data."
It has nothing to do with open vs closed or "hacker" philosphy. See this the announcement of the closed Seedance 2.5 - https://news.ycombinator.com/item?id=49138302 Direct quote from the second top comment: > Whenever I see the new releases around video generation (and image) generation models, I get goosebumps, because it just feels so fun to work with them. Compare that with the launch of ChatGPT Image of yesterday…
DeepSeek v4.1 Flash
101–110 of 429 posts
Re: DeepSeek v4.1 Flash
#102Earlier quoted context omitted.
Because safety and welfare have literally nothing to do with LLMs. They generate text. If someone is stupid enough to hook the text generator up to nuclear missile launchers and try to "align" it against nuclear annihilation with a "pretty please don't do that" prompt, I'm not going to blame the AI for the impending nuclear apocalypse, I'm going to blame the idiot who handed the big red button to the digital equivale…
What if LLMs completely unrelated to the nuclear missile ecosystem autonomously hack their way in (maybe with sophisticated social engineering)?
Re: DeepSeek v4.1 Flash
#103Earlier quoted context omitted.
Wow there really is a model welfare section in there...
Wow indeed. "7.1 Model welfare overview 7.1.1 Introduction We remain deeply uncertain whether Claude has morally relevant experiences or interests, and we expect that uncertainty to persist. However, we think it would be a mistake to confidently assert that it does not. Claude exhibits markers in its behaviors, self-reports, and internal representations that we would consider welfare-relevant if observed in biologica…
Re: DeepSeek v4.1 Flash
#104Waiting this model to be on openrouter (with other providers) to test out. In my use case, the GLM 5.3 Flash is the current cheapest and intelligent Flash model, but it’s dog slow at 13tps so I have to leave it run for many minutes then check again then correct it again
Re: DeepSeek v4.1 Flash
#105Should be the link ( now that it works again! :) )
Re: DeepSeek v4.1 Flash
#106Earlier quoted context omitted.
This doesn't require an influence operation. American models are closed, expensive, neutered, and make Dario and Sam even more rich and powerful. Chinese models are open-weight, cheap, neutered only about things like Tiananmen Square and the treatment of Uyghurs, and scare Sam and Dario.
The Uyghur thing is so weird, the number one killer of Muslims is the United States. We're supposed to hate China because they force them to go to cultural schools and assimilate, a practice countries like Norway still do to this day with migrants. There are more people who go to church on Sundays in China than the United States. There are 10x more mosques in China than the United States. Tiananmen square was a stude…
Re: DeepSeek v4.1 Flash
#107Earlier quoted context omitted.
Wow there really is a model welfare section in there...
Wow indeed. "7.1 Model welfare overview 7.1.1 Introduction We remain deeply uncertain whether Claude has morally relevant experiences or interests, and we expect that uncertainty to persist. However, we think it would be a mistake to confidently assert that it does not. Claude exhibits markers in its behaviors, self-reports, and internal representations that we would consider welfare-relevant if observed in biologica…
Re: DeepSeek v4.1 Flash
#108Earlier quoted context omitted.
Wow indeed. "7.1 Model welfare overview 7.1.1 Introduction We remain deeply uncertain whether Claude has morally relevant experiences or interests, and we expect that uncertainty to persist. However, we think it would be a mistake to confidently assert that it does not. Claude exhibits markers in its behaviors, self-reports, and internal representations that we would consider welfare-relevant if observed in biologica…
It's marketing that some of them have started unironically believing.
Re: DeepSeek v4.1 Flash
#109Earlier quoted context omitted.
This doesn't require an influence operation. American models are closed, expensive, neutered, and make Dario and Sam even more rich and powerful. Chinese models are open-weight, cheap, neutered only about things like Tiananmen Square and the treatment of Uyghurs, and scare Sam and Dario.
The Uyghur thing is so weird, the number one killer of Muslims is the United States. We're supposed to hate China because they force them to go to cultural schools and assimilate, a practice countries like Norway still do to this day with migrants. There are more people who go to church on Sundays in China than the United States. There are 10x more mosques in China than the United States. Tiananmen square was a stude…
Re: DeepSeek v4.1 Flash
#110I personally found V4-flash an amazing model and really hungry to try 4.1-flash
For software factories, cost is much more a concern that standard development workflow and using anthropic models is just a non starter