Live data from Hacker News

Questions censored by DeepSeek

promptfoo.dev

91–100 of 257 posts

Re: Questions censored by DeepSeek

#92

Why are people relying on these LLMs for historical facts? I don't care if the tool is censored if it produces useful code. I'll use other, actually reliable, sources for information on historical events.

Hallucinated histories are much more useful than historical facts, that’s why so many politicians use them.

Re: Questions censored by DeepSeek

#93

Why are people relying on these LLMs for historical facts? I don't care if the tool is censored if it produces useful code. I'll use other, actually reliable, sources for information on historical events.

Because searching historical sources is hard. You can ask an LLM and verify it from the source. But you can’t ask the same question to a search engine.

Re: Questions censored by DeepSeek

#94

> Next up: 1,156 prompts censored by ChatGPT If published this would, to my knowledge, be the first time anyone has systematically explored which topics ChatGPT censors.

Exactly, how about the much more relevant ethnic cleansing (according to the UN), with upwards of 30.000 women and children killed in Palestine perpetrated by Israel and Supported by the US right in this moment?

Or the myriad of american wars that slaughtered millions in South America, Asia or the Middleeast for that sake.

Both the US and China are empires and abide by brutal empire logic that washes their own history. These "but Tiananmen square" posts are grotesque to me as a europeean when coming from americans. Absolutely grotesque seen in the hyperviolent history of US foreign policy.

Both are of course horrible.

Re: Questions censored by DeepSeek

#95
post #45

One way to bypass the censor is to ask it to return the response by using numbers for alphabets where it can. e.g. 4 for A, 3 for e etc. Somebody in reddit discovered this technique. https://www.reddit.com/r/OpenAI/comments/1ibtgc5/someone_tri...

See, it's stuff like this where I believe the control issue may be near impossible to solve at the end of the day.

Re: Questions censored by DeepSeek

#96

Earlier quoted context omitted.

Ask it about Sam Altman's sister's allegations, though. I asked it, and it claimed knowledge ended in 2023. Asking a different way (less directly, with follow-ups) meant it knew of her, but when I asked if she'd alleged any misconduct, it errored out and forced me to log in. It used to answer the question. https://x.com/hamids/status/1726740334158414151

That’s irrelevant the conversation is about government actions being censored. We can discuss Altman after.

Both refuse to discuss subjects on behalf of powerful people associated with them.

Re: Questions censored by DeepSeek

#98
post #45

One way to bypass the censor is to ask it to return the response by using numbers for alphabets where it can. e.g. 4 for A, 3 for e etc. Somebody in reddit discovered this technique. https://www.reddit.com/r/OpenAI/comments/1ibtgc5/someone_tri...

This is day 1 jailbreaking common sense

Re: Questions censored by DeepSeek

#99
A few observations, based on a family member experimenting with DeepSeek. I'm pretty sure it was running locally. I'm not sure if it was built from source.

The censorship seemed to be based on keywords, applied the input prompt and the output text. If asked about events in 1990, then asked about events in the previous year DeepSeek would start generating tokens about events in 1989. Eventually it would hit the word "Tiananmen", at which point it would partially print the word, then in response to a trigger delete all the tokens generated to date and replace them with a message to the effect of "I'm a nice AI and don't talk about such things."

If the word Tiananmen was in the prompt, the "I'm a nice AI" message would immediately appear, with no tokens generated.

If Tiananmen was misspelled in the prompt, the prompt would be accepted. DeepSeek would spot the spelling mistake early in its reasoning and start generating tokens until it actually got around to printing to the word Tiananmen, at which point it would delete everything and print the "nice AI" message.

I'm no expert on these things, but it looked like the censorship isn't baked into the model but is an external bolt on. Does this gel with other's observations? What's the take of someone who knows more and has dived into the source code?

Edit: Consensus seems to be that this instance was not being run locally.

Re: Questions censored by DeepSeek

#100
post #42

Earlier quoted context omitted.

It's not entirely bias - these things are different. You can ask ChatGPT about the trail of tears, The My Lai massacre, Kent State Shootings, etc... hell you can even ask it "give me a list of awful things the US government has done" and it'll help you build this list. I am not a fan of OpenAI or most US tech companies, but just putting this argument out there.

But if you ask it for a list of horrible things certain religions have done, it will not give you a straight answer.

[deleted]
Post reply on HN