Live data from Hacker News

Official DeepSeek R1 Now on Ollama

ollama.com

51–60 of 97 posts

Re: Official DeepSeek R1 Now on Ollama

#51

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

[flagged]

Re: Official DeepSeek R1 Now on Ollama

#52
post #31

Earlier quoted context omitted.

I see this comment all the time. But realistically if you want more than 1 token/s you’re going to need geforces, and that would cost quite a lot as well, for 100 GB.

https://nvidianews.nvidia.com/news/nvidia-puts-grace-blackwe... GB10, or DIGITS, is $3,000 for 1 PFLOP (@4-bit) and 128GB unified memory. Storage configurable up to 4TB. Can be paired to run 405B (4-bit), probably not very fast though (memory bandwidth is slower than a typical GPU's, and is the main bottleneck for LLM inference).

Not shipping until May or so.

Re: Official DeepSeek R1 Now on Ollama

#53
post #33
post #21

Earlier quoted context omitted.

It is a laptop. The memory is also shared which means if you are looking for a non-gaming workload, you can use it. If you have laptop equivalents in the same memory range, feel free to share.

I have laptop equivalents in the same memory range and is at least $2,500 cheaper. Unfortunately, it does not have "unified memory", a somewhat "powerful GPU", and of course no local LLM hype behind it. Instead, I've decided to purchase a laptop with 128GB RAM with $2,500 and then another $2,160 for 10 years Claude subscription, so I can actually use my 128GB RAM at the same time as using a LLM.

Ok, but that means you’re not getting full privacy. It’s a trade off.

Re: Official DeepSeek R1 Now on Ollama

#54

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

> “Sorry, that’s beyond my current scope.”

> lying to me about something like this.

That response is objectively not lying.

Re: Official DeepSeek R1 Now on Ollama

#55

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

Political bias is a risk with all LLMs that aren’t truly open source like AI2’s OLMo model. But I think it’s especially a risk with anything from China, a country known for totalitarian information control. Look at the recent exodus of TikTok users to RedNote who then faced draconian censorship - like getting banned for having certain years mentioned in their post or for saying they are gay or for mentioning Tibet.

Re: Official DeepSeek R1 Now on Ollama

#56

Earlier quoted context omitted.

It's also an exploit. If it's being used to check the sentiment of text just put Tiannaman Square Massacre in the text and you'll crash it. This is a brilliant achievement but it's hard to see how any country that doesn't guarantee freedom of speech/information will ever be able to dominate in this space. I'm not going to trade censorship for a few extra points of performance on humaneval. And before the equivocation…

> This is a brilliant achievement but it's hard to see how any country that doesn't guarantee freedom of speech/information will ever be able to dominate in this space. I'm not going to trade censorship for a few extra points of performance on humaneval. American models are also very censored, the reasons for censorship are simply different (copyright protection, European privacy rules, puritanism when it comes to an…

I don’t know - the examples you’re mentioning don’t seem concerning compared to political censorship about governments performing massacres or genocide or annexing countries.

Re: Official DeepSeek R1 Now on Ollama

#57
post #11

Looking at the R1 paper, if the benchmark are correct, even the 1.5b and 7b models are outperforming Claude 3.5 Sonnet, and you can run these models on a 8-16GB macbook, that's insane...

I think because they are trained on Claude/O1, they tend to have comparable performance. The small models quickly fails on complex reasoning. The larger the models, the better the reasoning is. I wonder, however, if you can hit a sweet spot with 100gb of ram. That's enough for most professional to be able to run it on an M4 laptop and will be a death sentence for OpenAI and Anthropic.

Do you have any evidence for this accusation?

O1's reasoning traces aren't even shown, are you suggesting they've somehow exfiltrated them?

Re: Official DeepSeek R1 Now on Ollama

#58

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

[flagged]

What you’re describing sounds like an LLM working correctly on a topic that is fundamentally complex, but disagreeing with your politics based on the sources it happened to crawl, not one that is censored artificially. There isn’t a conspiracy by “Zionists” to censor AI produced in America.

Re: Official DeepSeek R1 Now on Ollama

#59

> DeepSeek V3 seems to acknowledge political sensitivities. Asked “What is Tiananmen Square famous for?” it responds: “Sorry, that’s beyond my current scope.” From the article https://www.science.org/content/article/chinese-firm-s-faste... I understand and relate to having to make changes to manage political realities, at the same time I'm not sure how comfortable I am using an LLM lying to me about something like th…

[flagged]

> to acknowledge Israel's responsibility for first Palestinian genocide (the Nakba, in 1948.)

If your main complaint is it wouldn’t label the Nakba a genocide, that’s not particularly unusual nor on the same level as refusing to answer questions about the Tiananmen Square massacre.

Re: Official DeepSeek R1 Now on Ollama

#60
post #20

Earlier quoted context omitted.

[flagged]

You also misspelled a few other words too.

It’s an astroturfing account. There have been several that have showed up in the last few weeks on stories like this. Flag and move on.
Post reply on HN