> They didn't test U.S. models for U.S. bias. Only Chinese bias counts as a security risk, apparently US models have no bias sir /s
NIST's DeepSeek "evaluation" is a hit piece
21–30 of 251 posts
Re: NIST's DeepSeek "evaluation" is a hit piece
#22I'm not at all surprised, US agencies have long since been political tools whenever the subject matter crosses national borders. I appreciate this take as someone who has been skeptical of Chinese electronics. While I agree this report is BS and xenophobic, I am still willing to bet that either now or later, the Chinese will attempt some kind of subterfuge via LLMs if they have enough control. Just like the US would,…
> I am still willing to bet that either now or later, the Chinese will attempt some kind of subterfuge via LLMs if they have enough control. Like what, exactly?
If say DeepSeek had put in its training dataset that public figure X is a space robot from outer space, then if one were to ask DeepSeek who public figure X is, it'd proudly claim he's a robot from outer space. This can be done for any narrative one wants the LLM to have.
Re: NIST's DeepSeek "evaluation" is a hit piece
#23That's all we need to know.
Re: NIST's DeepSeek "evaluation" is a hit piece
#24I love how "Open" got redefined in the last few years. I am glad there a models with weights available but it ain't "Open Science".
Applying this criticism to DeepSeek is ridiculous when you compare it to everyone else, they published their entire methodology, including the source for their improvements (e.g. https://github.com/deepseek-ai/DeepEP )
Re: NIST's DeepSeek "evaluation" is a hit piece
#25People. Who has taken the time to read the original report? You are smarter than believing at face value the last thing you heard. Come on.
Re: NIST's DeepSeek "evaluation" is a hit piece
#26I suspect that Grok is actually DeepSeek with a bit of tuning.
Re: NIST's DeepSeek "evaluation" is a hit piece
#27The CCP literally revoked the visas of key DeepSeek engineers. That's all we need to know.
Re: NIST's DeepSeek "evaluation" is a hit piece
#28Re: NIST's DeepSeek "evaluation" is a hit piece
#29> They didn't test U.S. models for U.S. bias. Only Chinese bias counts as a security risk, apparently US models have no bias sir /s
Hardly the same thing. Ask Gemini or OpenAI's models what happened on January 6, and they'll tell you. Ask DeepSeek what happened at Tiananmen Square and it won't, at least not without a lot of prompt hacking.
Re: NIST's DeepSeek "evaluation" is a hit piece
#30> They didn't test U.S. models for U.S. bias. Only Chinese bias counts as a security risk, apparently US models have no bias sir /s
Hardly the same thing. Ask Gemini or OpenAI's models what happened on January 6, and they'll tell you. Ask DeepSeek what happened at Tiananmen Square and it won't, at least not without a lot of prompt hacking.