Live data from Hacker News

Promising results from DeepSeek R1 for code

simonwillison.net

481–490 of 765 posts

Re: Promising results from DeepSeek R1 for code

#481

Earlier quoted context omitted.

yeah: aider --model ollama_chat/deepseek-r1:32b (or whatever)

This didn't work well for me, no changes are ever made but maybe it's because I'm just using the 14B model.

In case you are on a 32+GB Mac, you could try deepseek-r1-distill-qwen-32b-mlx in LM Studio. It’s just barely usable speed-wise, but gives useful results most of the time.

Re: Promising results from DeepSeek R1 for code

#482

Earlier quoted context omitted.

If you believe that everything will be solved by the time you can pivot, what will we need jobs for anyway? I mean, the bottleneck justifying most scarcity is that we don't have adequate software to ask the robots to do the thing, so if that's a solved problem, which things will remain that still need doing? I don't personally think that's how it will go. AI will always need its hand held, if not due to a lack of cap…

I'm a student, so all pivots have a minimum delta of 2 years, which is something like a 100x on current capabilities on the seemingly steep s-curve we are on. That drives my "gloom" (in practice I've placed my hope in something eternal rather than a fickle thing like this)

What he meant is that if this really happens, and LLMs replaces humans everywhere and everybody becomes unemployed, congratulations you'll be fine.

Because at that point there's 2 scenarios:

- LLMs don't need humans anymore and we're either all dead or in a matrix-like farm

- Or companies realize they can't make LLMs buy the stuff their company is selling (with what money??) so they still need people to have disposable income and they enact some kind of Universal Basic Income. You can spend your days painting or volunteering at an animal shelter

Some people are rooting for the first option though, so while it's good that you've found faith, another thing that young people are historically good at is activism.

Re: Promising results from DeepSeek R1 for code

#483

Earlier quoted context omitted.

In the past, human workers were displaced. The value of their labour for certain tasks became lower than what automation could achieve, but they could still find other things to do to earn a living. What people are worrying about here is what happens when the value of human labour drops to zero, full stop. If AI becomes better to us at everything, then we will do nothing, we will earn nothing, and we will have nothin…

If anything like that had actually happened in the past, you might have a point. When it comes to what happens when the value of human labor drops to zero, my guess is every bit as good as yours. I say it will be a Good Thing. "Work" is what you call whatever you're doing when you'd rather be doing something else.

The value of our labour is what enables us to acquire things and property, with which we can live and do stuff. If your labour is valueless because robots can do anything you can do better, how do you get any of the possessions you require in order to do that something else you'd rather be doing? Capitalism won't just give them to you. If you do not own land, physical resources or robots, and you can't work, how do you get food? Charity? I'd argue there will need to be a pretty comprehensive redistribution scheme for the people at large to benefit.

Re: Promising results from DeepSeek R1 for code

#484

Earlier quoted context omitted.

Noob question (I only learned how to use ollama a few days ago): what is the easiest way to run this DeepSeek-R1-Distill-Qwen-32B model that is not listed on ollama (or any other non-listed model) on my computer ?

This model is listed on ollama. The 20GB one is this one: https://ollama.com/library/deepseek-r1:32b-qwen-distill-q4_K...

Ok, the "View all" option in the dropdown is what I missed! Thanks!

Re: Promising results from DeepSeek R1 for code

#485
post #395

Earlier quoted context omitted.

Noob question (I only learned how to use ollama a few days ago): what is the easiest way to run this DeepSeek-R1-Distill-Qwen-32B model that is not listed on ollama (or any other non-listed model) on my computer ?

Search for a GGUF on Hugging Face and look for a "use this model" menu, then click the Ollama option and it should give you something to copy and paste that looks like this: ollama run hf.co/MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF:IQ1_M

Got it, thank you!

Re: Promising results from DeepSeek R1 for code

#486

Earlier quoted context omitted.

I'm a student, so all pivots have a minimum delta of 2 years, which is something like a 100x on current capabilities on the seemingly steep s-curve we are on. That drives my "gloom" (in practice I've placed my hope in something eternal rather than a fickle thing like this)

What he meant is that if this really happens, and LLMs replaces humans everywhere and everybody becomes unemployed, congratulations you'll be fine. Because at that point there's 2 scenarios: - LLMs don't need humans anymore and we're either all dead or in a matrix-like farm - Or companies realize they can't make LLMs buy the stuff their company is selling (with what money??) so they still need people to have disposab…

The scenario that is worrying is having to deal with the jagged frontier of intelligence prolonging the hurt. i.e

202X: SWE is solved

202X + Y; YIn this case, I can't retrain before the second threshold but also can't idle. I just have to suffer. I'm prepared to, but it's hard to escape fleshy despair.

Re: Promising results from DeepSeek R1 for code

#487
post #389

Earlier quoted context omitted.

Why do people keep talking about this? We get it, Chinese models are censored by CCP law. Can we stop talking about it now? I swear this must be some sort of psyop at this point.

> Can we stop talking about it now? I swear this must be some sort of psyop at this point. It's not a psyop that people in democracies want freedom. Democrats (not the US party) know that democracy is fragile. That's why it's called an "experiment". They know they have to be vigilant. In ancient Rome it was legal to kill on the spot any man who attempted to make himself king, and the Roman Republic still fell. Many p…

Don't worry, the way things are going, you'll have that in the US as well soon.

Ironically supported by the folks who argue that having an assault rifle at home is an important right to prevent the government from misusing its power.

Re: Promising results from DeepSeek R1 for code

#488
post #389

Earlier quoted context omitted.

Why do people keep talking about this? We get it, Chinese models are censored by CCP law. Can we stop talking about it now? I swear this must be some sort of psyop at this point.

Mostly anti-Chinese bias from Americans, Western Europeans, and people aligned with that axis of power (e.g. Japan). However, on the Japanese internet, I don't see this obsession with taboo Chinese topics like on Hacker News. People on Hacker News will rave about 天安門事件 but they will never have heard of the South Korean equivalent (cf. 光州事件) which was supported by the United States government. I try to avoid discussin…

The exact same discussions were going on with "western" models. Don't remember the images of black nazis making the rounds because inclusion? Same thing. This HN tread is the first time I'm hearing about this anti-DeepSeek sentiment, so arguably it's on a lower level actually.

So let's not get too worked up, shall we?

Re: Promising results from DeepSeek R1 for code

#489
post #88

Earlier quoted context omitted.

Any person who has the ability to break down a problem to the point that code can be written to solve it, and the ability to work with an LLM system to get that work done, and the ability to evaluate if the resulting code solves the problem. That's a mixture of software developer, program manager, product manager and QA engineer. I think that's what software developer roles will look like in the future: a slightly di…

I really want this to be true, but honestly it's really hard. What makes you think this won't be eaten too within the next year based on the current s-curve-if-not-exponential we are on?

I still don't believe in AGI.

Re: Promising results from DeepSeek R1 for code

#490

Earlier quoted context omitted.

Mostly anti-Chinese bias from Americans, Western Europeans, and people aligned with that axis of power (e.g. Japan). However, on the Japanese internet, I don't see this obsession with taboo Chinese topics like on Hacker News. People on Hacker News will rave about 天安門事件 but they will never have heard of the South Korean equivalent (cf. 光州事件) which was supported by the United States government. I try to avoid discussin…

The exact same discussions were going on with "western" models. Don't remember the images of black nazis making the rounds because inclusion? Same thing. This HN tread is the first time I'm hearing about this anti-DeepSeek sentiment, so arguably it's on a lower level actually. So let's not get too worked up, shall we?

The black nazis thing wasn't caused by government regulation of models.
Post reply on HN