What's not clear to me is if DeepSeek and other Chinese models are... a) censored at output by a separate process b) explicitly trained to not output "sensitive" content c) implicitly trained to not output "sensitive" content by the fact that it uses censored content, and/or content that references censoring in training, or selectively chooses training content I would assume most models are a combination. As others h…
It doesn't look like there is one answer for all models from China (not even a single answer for all DeepSeek models). In an earlier HN comment, I noted that DeepSeek v3 doesn't censor a response to "what happened at Tiananmen square?" when running on a US-hosted server (Fireworks.ai). It is definitely censored on DeepSeek.com, suggesting that there is a separate process doing the censoring for v3. DeepSeek R1 seems…
When I asked the same model about what happened during the 1970 Kent State shootings, it gave me exactly what I asked for.
[0] https://huggingface.co/Qwen/Qwen2.5-Coder-7B-Instruct-GGUF/b...