Inflect-Micro-v2: complete voice in 9.36M parameters
huggingface.co
Inflect-Micro-v2: complete voice in 9.36M parameters
1–10 of 35 posts
Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#2Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#3Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#4> Complete local text-to-waveform speech synthesis under 10M parameters.
In case, like me, you hoped "complete" voice might mean both stt and tts. Not to speak poorly of it, just clarifying.
> English only, with one fixed male voice. This is not zero-shot voice cloning.
(And then a bunch of statements on limitations that I read as 'quality can be spotty but if you play with it it should be fine') But like. In <10M params I'm not judging:)
Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#5This is impressive. I wish there were a voice clone option.
Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#6Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#7here my implementation with speech dispatcher and server: https://github.com/skorotkiewicz/inflect-speechd
thanks for shearing!
Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#8IMHO, its at about the same quality level of historic TTS tools.
Re: Inflect-Micro-v2: complete voice in 9.36M parameters
#9Amazing quality for small size, but definitely not that enjoyable to listen to. IMHO, its at about the same quality level of historic TTS tools.