DeepSeek R1: Open Weights, Hidden Bias
blog.getplum.ai
DeepSeek R1: Open Weights, Hidden Bias
1–9 of 9 posts
Re: DeepSeek R1: Open Weights, Hidden Bias
#2We evaluated DeepSeek R1 and confirmed that its guardrails deviate significantly from other model providers. We’re currently updating it to behave more in line with Anthropic and OpenAI’s models.
Re: DeepSeek R1: Open Weights, Hidden Bias
#3Re: DeepSeek R1: Open Weights, Hidden Bias
#4Re: DeepSeek R1: Open Weights, Hidden Bias
#5The bias is baked into the open weights, namely happening on self-hosted 671B LLM??
Re: DeepSeek R1: Open Weights, Hidden Bias
#6The bias is baked into the open weights, namely happening on self-hosted 671B LLM??
Yes -- we observed this behavior on both the open-source open-weights 671B model as well as the DeepSeek web app.
Then I have mixed signals about this.
Re: DeepSeek R1: Open Weights, Hidden Bias
#7Earlier quoted context omitted.
Yes -- we observed this behavior on both the open-source open-weights 671B model as well as the DeepSeek web app.
Weird, because I got some deepseek feedback where it was openly critical and explicit about the authoritative regime of china. I really thought it was the "deepseek web app" only. Then I have mixed signals about this.
Re: DeepSeek R1: Open Weights, Hidden Bias
#8Earlier quoted context omitted.
Weird, because I got some deepseek feedback where it was openly critical and explicit about the authoritative regime of china. I really thought it was the "deepseek web app" only. Then I have mixed signals about this.
We're working on a follow-up post focused on our analysis of the open-source open-weight 671B model. What we're seeing is that questions related to the Chinese government produce an empty chain-of-thought followed by pro-Chinese-government talking points.
This is going to be very hard to trust anything about it anymore, unless running the 671B locally an my own systems.
Re: DeepSeek R1: Open Weights, Hidden Bias
#9Earlier quoted context omitted.
We're working on a follow-up post focused on our analysis of the open-source open-weight 671B model. What we're seeing is that questions related to the Chinese government produce an empty chain-of-thought followed by pro-Chinese-government talking points.
It is too late, I got mixed signals. This is going to be very hard to trust anything about it anymore, unless running the 671B locally an my own systems.
Happy to send you the dataset if you'd like! Please reach out to our email linked in the post.