Live data from Hacker News

Questions censored by DeepSeek

promptfoo.dev

251–257 of 257 posts

Re: Questions censored by DeepSeek

#251

Earlier quoted context omitted.

Groq doesn’t have r1, only a llama 70b distilled with r1 outputs. Kinda crazy how they just advertise it as actual r1

I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this?

Actual r1 is a base model by Deepseek with CoT added by them via RL.

The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.

As I understand it.

Re: Questions censored by DeepSeek

#252
post #213

Earlier quoted context omitted.

The Deepseek hosted chat site has additional 'post-hoc' censorship applied from what people have observed, if that's what you're referring to. While the foundational model (including self hosted) has some just part of its training which is the kind the article is discussing, yes.

Thanks for cutting through the noise. I did some poking around and a discussion from a couple of days ago reached the same conclusion. https://news.ycombinator.com/item?id=42825573

Is it correct or incorrect that they open-sourced tbeir code? i.e. can anyone with $6M now take the DeepSeek training code, apply it to their dataset of interest, and train a new model that is not censoeed (i.e. even somehow intrinsically to the kodel itself)? Apologies I am not an AI engineer nor even a software engineer of my terminology usage isn't quite spot on.

Re: Questions censored by DeepSeek

#253
post #213

Earlier quoted context omitted.

Thanks for cutting through the noise. I did some poking around and a discussion from a couple of days ago reached the same conclusion. https://news.ycombinator.com/item?id=42825573

Is it correct or incorrect that they open-sourced tbeir code? i.e. can anyone with $6M now take the DeepSeek training code, apply it to their dataset of interest, and train a new model that is not censoeed (i.e. even somehow intrinsically to the kodel itself)? Apologies I am not an AI engineer nor even a software engineer of my terminology usage isn't quite spot on.

They have definitely open sourced the inference code. I haven't any training code. I don't think HAI-LLM is open source.

But certainly you can take the architecture from the paper and train a similar model. Or you can try to remove the alignment and produce and uncensored version then realign it.

But at least part of the advantage they have is training on Chinese internet data from inside the great firewall that (AFAIK) US companies don't have access to for any price.

Re: Questions censored by DeepSeek

#254
post #218
post #128

Earlier quoted context omitted.

I ran the 32b parameter model just fine on my rig an hour ago with a 4090 and 64gig of ram. It’s high end for the consumer scene but still solidly within consumer prices

I run the 32b parameter model also just fine on our 4x H100 rig :) It's good enough for embedding, our use-case.

I'm not sure if $200k of hardware fits the consumer level

Re: Questions censored by DeepSeek

#255

It's interesting to see the number of comments that consist of whataboutism ("But, but, but ChatGPT!") and minimization of the problem ("It's not really censorship." or "You can get that information elsewhere."). I like to avoid conspiracy theories, but it wouldn't surprise me if the CCP were trying to make DeepSeek more socially acceptable.

Yeah same. Even in this very thread there were thankfully flagged commenters that were pro-China sockpuppets.

In the Wikipedia article for whataboutism, one can see that such tactics were a large mainstay of Soviet Union propaganda.

Re: Questions censored by DeepSeek

#256

Earlier quoted context omitted.

Exactly, how about the much more relevant ethnic cleansing (according to the UN), with upwards of 30.000 women and children killed in Palestine perpetrated by Israel and Supported by the US right in this moment? Or the myriad of american wars that slaughtered millions in South America, Asia or the Middleeast for that sake. Both the US and China are empires and abide by brutal empire logic that washes their own histor…

Virtually all countries within the European continent have been perpetrators of colonialism and genocide in the past 4 centuries, several in the last 90 years, and a few in the last 20 years. It is a banal observation. The reason why the string "tiananmen" is so frequently invoked is that it is a convenient litmus test for censorship/alignment/whatever-your-preferred-term by applications that must meet Chinese govern…

[deleted]

Re: Questions censored by DeepSeek

#257
post #142
post #128

Earlier quoted context omitted.

I ran the 32b parameter model just fine on my rig an hour ago with a 4090 and 64gig of ram. It’s high end for the consumer scene but still solidly within consumer prices

I have also been running the 32b version on my 24GB RTX 3090.

That's not DeepSeek, it's a Qwen or Llama model distilled from DeepSeek. Not the same thing at all.
Post reply on HN