Untitled topic
1–10 of 21 posts
Re: undefined
#2The model is Apache 2.0 licensed and trained on less than 100 hours of audio data. It supports both American and British English, offering multiple voice options with natural emotional expression and 24kHz audio output.
We've deployed a demo at kokorotts.online where you can try it out. I'd really appreciate any feedback from the HN community on both the model's performance and potential applications.
Tech stack: StyleTTS 2 architecture, ONNX runtime, Next.js for the web interface.
Re: undefined
#3it is very fast and very passable.
Re: undefined
#4I'm excited to share Kokoro TTS, an open-source text-to-speech model we've been working on. Despite its relatively small size (82M parameters), it achieves impressive results in natural speech synthesis, ranking first in the TTS Spaces Arena benchmark. The model is Apache 2.0 licensed and trained on less than 100 hours of audio data. It supports both American and British English, offering multiple voice options with…
Re: undefined
#5I'm excited to share Kokoro TTS, an open-source text-to-speech model we've been working on. Despite its relatively small size (82M parameters), it achieves impressive results in natural speech synthesis, ranking first in the TTS Spaces Arena benchmark. The model is Apache 2.0 licensed and trained on less than 100 hours of audio data. It supports both American and British English, offering multiple voice options with…
It's NOT Open Source.
Re: undefined
#6I'm excited to share Kokoro TTS, an open-source text-to-speech model we've been working on. Despite its relatively small size (82M parameters), it achieves impressive results in natural speech synthesis, ranking first in the TTS Spaces Arena benchmark. The model is Apache 2.0 licensed and trained on less than 100 hours of audio data. It supports both American and British English, offering multiple voice options with…
It's NOT Open Source.
"There currently isn't a release date scheduled for the other voices"
[1]: https://huggingface.co/blog/hexgrad/kokoro-short-burst-upgra...
Re: undefined
#7I'm excited to share Kokoro TTS, an open-source text-to-speech model we've been working on. Despite its relatively small size (82M parameters), it achieves impressive results in natural speech synthesis, ranking first in the TTS Spaces Arena benchmark. The model is Apache 2.0 licensed and trained on less than 100 hours of audio data. It supports both American and British English, offering multiple voice options with…
Re: undefined
#8I'm excited to share Kokoro TTS, an open-source text-to-speech model we've been working on. Despite its relatively small size (82M parameters), it achieves impressive results in natural speech synthesis, ranking first in the TTS Spaces Arena benchmark. The model is Apache 2.0 licensed and trained on less than 100 hours of audio data. It supports both American and British English, offering multiple voice options with…
It's NOT Open Source.
- Apache 2.0 weights in this repository
- MIT inference code in spaces/hexgrad/Kokoro-TTS adapted from yl4579/StyleTTS2
- GPLv3 dependency in espeak-ng
Re: undefined
#9> Can I use Kokoro TTS offline?
> Kokoro TTS is a cloud-based service that requires an internet connection to access our advanced text to speech technology. This ensures you always have access to the latest improvements and don't need to worry about local hardware requirements or model installations.
I would happily take on the worrying for offline instead of them having to worry about my worries.
Re: undefined
#10And in the FAQ:
> What's included in the Kokoro TTS free trial?
> New users can try Kokoro TTS's full capabilities with our free trial. This allows you to experience our professional-grade text to speech technology firsthand, including access to all voices and both American and British English options.
So this is the "free trial"? Plus it being a cloud-based service makes me not understand the situation.