Live data from Hacker News

VibeVoice: Open-source frontier voice AI

github.com

1–10 of 191 posts

Re: VibeVoice: Open-source frontier voice AI

#8
Seems quite heavy for a STT model, Parakeet and Whisper are much smaller and perform great for quick dictation and transcription of longer files. I guess that's due to additional accuracy and speaker diarisation?

The TTS example clip in the repo of 'spontaneous singing' is creepy as fuck

Re: VibeVoice: Open-source frontier voice AI

#9

Isn't this project the one Microsoft published but then soon after pulled it for security/safety reasons? What has changed since then?

Look at the "News" section in the readme - The original TTS model is gone from this repo (you can still find it other places), but the SST/ASR, long form TTS, and streaming TTS models are newer.
Post reply on HN