Okay but does any one actually _want_ a reasoning model at such low tok/sec speeds?!
Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
51–60 of 162 posts
Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#52Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#53Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#54Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#55Earlier quoted context omitted.
It is shown running on 2 or 4 raspberry pis; the point is that you can add more (ordinary, non GPU) hardware for faster inference. It's a distributed system. The sky is the limit.
ah, Thanks! but what can a distributed system like this do? is this a fun to do, for the sake of doing it project or does it have practical applications? just curious about applicability thats all.
The advantage over centralizing the compute is that you can just connect your node and start contributing to the cause (both by providing compute and by being its eyes and hands out there in the real world), there's no confusion over things like who is paying the cloud compute bill and nobody has invested overmuch in hardware.
Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#56When can I "apt-get install" all this fancy new AI stuff?
Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#57Earlier quoted context omitted.
I really don't like that these models can be branded as Deepseek R1.
Well, Deepseek trained them?
Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#58That's not a bad result, although for £320 for 4x Pi5s you could probably find a used 12GB 3080 and probably more than 10x token speed
That wouldn't get on Hacker News ;-)
Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#59Re: Deepseek R1 Distill 8B Q40 on 4 x Raspberry Pi 5
#60Does adding memory help? There's a Rpi 5 with 16GB RAM recently available.