Earlier quoted context omitted.
It should participate in all conversations, take initiative and experiment.
"Hey, hey, are you still asleep? Using spare cycles, I have designed an optimal recipe for mashed potatoes, as you mentioned ten days ago. I need you to go get some potatoes."
A 30B Qwen model walks into a Raspberry Pi and runs in real time
111–120 of 139 posts
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#112Earlier quoted context omitted.
"Hey, hey, are you still asleep? Using spare cycles, I have designed an optimal recipe for mashed potatoes, as you mentioned ten days ago. I need you to go get some potatoes."
I unironically want this.
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#113Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#114Earlier quoted context omitted.
"Hey, hey, are you still asleep? Using spare cycles, I have designed an optimal recipe for mashed potatoes, as you mentioned ten days ago. I need you to go get some potatoes."
A local AI system that hears your conversations, identifies problems, and then uses spare cycles to devise solutions for them is actually an incredible idea. I'm never going to give a cloud system the kind of access it would need to do a really good job, but a local one I control? Absolutely. "Hey, are you still having trouble with[succinct summary of a problem it identified]?" "Yes" "I have a solution that meets you…
I call that Dreaming.
(TM)
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#115How can I use ByteShape to run LLMs faster on my 32GB MacBook M1 Max? Or has Ollama already optimized that?
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#116Earlier quoted context omitted.
I’ve experimented with several of the really small models. It’s impressive that they can produce anything at all, but in my experience the output is basically useless for anything of value.
Yes, I thought that too! But qwen3:0.6b (and to some extent gemma 1b) has made me reevaluate. They still aren't useful like large LLMs, but for things like summarization, and other tasks where you can give them structure but want the sheen of natural language they are much better than things like the Phi series were.
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#117Earlier quoted context omitted.
> but I've have great success with Home Assistant and the Home Assistant Voice Preview edition As compared to Alexa? I bought their preview hardware (and had a home-rolled ESP32 version before that even) and things are getting closer, I can see the future where this works but we aren't there today IMHO. HA Voice (the current hardware) does not do well enough in the mic or speaker [0] department when compared to the E…
I've not had any issues with the audio picking up, but its in the living room rather than the kitchen. I have Alexa's in most rooms. I don't play music through it, which I do from the Alexa. Tbh I think the mic and the speakers will be fine when the rest of the 'product' is sorted. I failed to mention I have Claude connected to it rather than their default assistant. To us, this just beats Alexa hands down. I have th…
I will say, while it was too slow (today) with the my local inference hardware (CPU, older computer and a little on my newer MBP) it was magically to talk to and hear back from HA all locally. I look forward to a future where I can do that at the same speed/quality as the cloud models. Yes, I know cloud models will continue to get better but turning on/off my fans/lights/etc doesn't need to best model available, just needs to be reliable and fast, I'm even fine with it "shelling out" to the cloud if I ask it for something outside of the basics though I doubt I'll care to do that.
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#118There is a huge market segment waiting here. At least I think there is. Well, at least people like me want this. Ok, tens of dollars can be made at least. It is just missing a critical tipping point. Basically, I want an alexa like device for the home backed by local inference and storage with some standardized components identified: - the interactive devices - all the alexa/google/apple devices out there are this in…
I've just started using it but I'd recommend https://github.com/steipete/clawdis , you need to set it up a bit but it's really cool to just be able to do things on the go by just texting an assistant. You can see all the different ways people are using it @clawdbot on twitter.
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#119Earlier quoted context omitted.
A local AI system that hears your conversations, identifies problems, and then uses spare cycles to devise solutions for them is actually an incredible idea. I'm never going to give a cloud system the kind of access it would need to do a really good job, but a local one I control? Absolutely. "Hey, are you still having trouble with[succinct summary of a problem it identified]?" "Yes" "I have a solution that meets you…
> A local AI system that hears your conversations, identifies problems, and then uses spare cycles to devise solutions for them is actually an incredible idea. I call that Dreaming. (TM)
Doubly so if you could just talk and brainstorm while it's listening and condensing, so you can circle back later and see what raindrops formed from the brainstorm.
Call that DayDreaming (TM)
Re: A 30B Qwen model walks into a Raspberry Pi and runs in real time
#120Original: 11tok/s Byteshape: 16tok/s
Quite a nice improvement!