It uses distributed models so latency is something I'm working on, but I am curious on where this could go, if anywhere.
Very much a WIP. Feedback welcome :-)
Show HN: Pi-C.A.R.D, a Raspberry Pi Voice Assistant
github.com
1–10 of 101 posts
It uses distributed models so latency is something I'm working on, but I am curious on where this could go, if anywhere.
Very much a WIP. Feedback welcome :-)
Show HN: Pi-C.A.R.D, a Raspberry Pi Voice Assistant
github.com
Missed opportunity for LCARS - LLM Camera Audio Recognition Service, responding to the keyword "computer," naturally. I guess if this ran elsewhere from a Pi, it could be LCARS.
> Why Pi-card? > Raspberry Pi - Camera Audio Recognition Device. Missed opportunity for LCARS - LLM Camera Audio Recognition Service, responding to the keyword "computer," naturally. I guess if this ran elsewhere from a Pi, it could be LCARS.
Props, and thank you for this.
> Why Pi-card? > Raspberry Pi - Camera Audio Recognition Device. Missed opportunity for LCARS - LLM Camera Audio Recognition Service, responding to the keyword "computer," naturally. I guess if this ran elsewhere from a Pi, it could be LCARS.
Pi-C.A.R.D is perfect. Read it 100% as Picard, and more recognizable that LCARS.
> Why Pi-card? > Raspberry Pi - Camera Audio Recognition Device. Missed opportunity for LCARS - LLM Camera Audio Recognition Service, responding to the keyword "computer," naturally. I guess if this ran elsewhere from a Pi, it could be LCARS.
LOCUTUS
> Why Pi-card? > Raspberry Pi - Camera Audio Recognition Device. Missed opportunity for LCARS - LLM Camera Audio Recognition Service, responding to the keyword "computer," naturally. I guess if this ran elsewhere from a Pi, it could be LCARS.
I see that a speaker is in the hardware list - does this speak back?
Additionally, since I'm streaming the LLM response, it won't take long to get your reply. Since it does it a chunk at a time, there's occasionally only parts of words that are said momentarily. Also of course depends on what model you use or what the context size is for how long you need to wait.
I wanted to create a voice assistant that is completely offline and doesn't require any internet connection. This is because I wanted to ensure that the user's privacy is protected and that the user's data is not being sent to any third party servers. Props, and thank you for this.