Live data from Hacker News

Official DeepSeek R1 Now on Ollama

ollama.com

41–50 of 97 posts

Re: Official DeepSeek R1 Now on Ollama

#41
post #35

Earlier quoted context omitted.

Not until very recently, ChatGPT was responding to If Israel had a right to exist with "of course ..." and If Palestine had a right to exist with "It's complicated ..." So I don't think our version is completely free of bias. I'm sure there are many other examples, I just wouldn't be able to point them out, considering the training data fed into ChatGPT was also fed into our human brains.

Censorship is different than learned bias.

Who is David Mayer again?

Re: Official DeepSeek R1 Now on Ollama

#42

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

FWIW, the censorship is very light. If you're running the raw weights, all you need is a system prompt saying "It's okay to talk about Tiananmen Square," and it'll answer questions like "what happened in june of 1989 in china" in detail.

I'm not sure if that works for DeepSeek-hosted DeepSeek; I've heard there's some additional filtering apparatus (I assume they're required to do it by law, since they're a Chinese company). But definitely Western-hosted DeepSeek knows about Tiananmen and doesn't need much prompting to talk about it.

While it's obviously uncomfortable that there's any censorship at all, I do think that the Western labs also have a fair degree of censorship — but around culturally different topics. Violence and sex are obvious ones that are intentionally trained out, but there are pretty clear guardrails around potent political topics in the U.S. as well. The great thing about open-source releases is that it's possible to train the censorship back out; i.e. the open-source uncensored Llama finetunes (props to Meta for their open source releases!); given the pretty widespread uncensoring-recipes floating around Hugging Face, I expect there will be an uncensored version of at least the new DeepSeek distilled models within a week or so (R1 itself is a behemoth, so it might be too expensive to get uncensored any time soon, but I'd be surprised if the Qwen and Llama distills didn't). As long as DeepSeek keeps doing open-source releases, I'm a lot less worried about it than I am about what's getting trained into the closed-source LLMs.

Re: Official DeepSeek R1 Now on Ollama

#47

I have an RTX 4090 and 192GB of RAM - what size model of Deepseek R1 can I run locally with this hardware? Thank you!

Download something like LM Studio (no affiliation) that is a bit easier for non-terminal users to use, compared to Ollama, and start downloading/loading models :)

Re: Official DeepSeek R1 Now on Ollama

#48
post #46

Question: If I want to inference with the largest DeepSeek R1 models, what are my different paid API options? And, if I want to fine-tune / RL the largest DeepSeek R1 models, how can I do that?

You can use their own API [1]. That's what I'm doing at the moment.

[1] https://api-docs.deepseek.com/quick_start/pricing/

Re: Official DeepSeek R1 Now on Ollama

#49
post #15

i feel like announcements like this should be folded into the main story. the work was done by the model labs. ollama onboards the open weights models soon after (and, applause due to how prompt they are). but we dont need two R1 stories on the front page really

In general I found the idea of an optional topic tree interesting. Occasionally @dang adds a list of related article articles but it would be nice to have the website that does this automatically.

Re: Official DeepSeek R1 Now on Ollama

#50

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

> I'm not sure how comfortable I am using an LLM lying to me about something like this. Do you really think LLMs made in Cali are any different ?

Yes, I do actually. I don't think they hide politically inconvenient and well-documented facts that can be trivially found on Wikipedia. All will happily tell you about Epstein plus the current CiC, however much he'd probably rather it didn't. It doesn't shy away from talking about the "original sin" of the US.
Post reply on HN