Live data from Hacker News

Deep-learning text-to-speech tool for generating voices of various characters

15.ai

81–88 of 88 posts

Re: Deep-learning text-to-speech tool for generating voices of various characters

#81

Are there open source projects like this?

There's Mozilla TTS https://github.com/mozilla/TTS

Here's a sample from a TTS model + vocoder I released for it. I've no wish to deter the motivated, but it'd take a bit of figuring out how to set things up and you'd need to read the docs and code to get oriented :)

https://m.soundcloud.com/user-726556259/sherlock-wavegrad-sa...

Links to the models are here: https://discourse.mozilla.org/t/creating-a-github-page-for-h...

Is originally trained on two novels read by the same narrator on LibriVox (ie in public domain)

Re: Deep-learning text-to-speech tool for generating voices of various characters

#83

Are there open source projects like this?

There's Mozilla TTS https://github.com/mozilla/TTS Here's a sample from a TTS model + vocoder I released for it. I've no wish to deter the motivated, but it'd take a bit of figuring out how to set things up and you'd need to read the docs and code to get oriented :) https://m.soundcloud.com/user-726556259/sherlock-wavegrad-sa... Links to the models are here: https://discourse.mozilla.org/t/creating-a-github-page-for-…

Is there a simple interface like the example in this thread to use the tool for a non developer regard Mozila TTS yet? I can't find one...

Re: Deep-learning text-to-speech tool for generating voices of various characters

#84

Earlier quoted context omitted.

There's Mozilla TTS https://github.com/mozilla/TTS Here's a sample from a TTS model + vocoder I released for it. I've no wish to deter the motivated, but it'd take a bit of figuring out how to set things up and you'd need to read the docs and code to get oriented :) https://m.soundcloud.com/user-726556259/sherlock-wavegrad-sa... Links to the models are here: https://discourse.mozilla.org/t/creating-a-github-page-for-…

Is there a simple interface like the example in this thread to use the tool for a non developer regard Mozila TTS yet? I can't find one...

There's the demo server which has a simple web UI where you can input text to be spoken, but in regards to setting it up locally it's not that suited for a non developer

https://github.com/mozilla/TTS/tree/master/TTS/server

https://github.com/mozilla/TTS/wiki/Build-instructions-for-s...

There's also a version in docker: https://github.com/synesthesiam/docker-mozillatts

And various Colabs too, which are fairly easy to get going with: https://github.com/mozilla/TTS/wiki/TTS-Notebooks-and-Tutori...

Re: Deep-learning text-to-speech tool for generating voices of various characters

#85

Are there open source projects like this?

There's Mozilla TTS https://github.com/mozilla/TTS Here's a sample from a TTS model + vocoder I released for it. I've no wish to deter the motivated, but it'd take a bit of figuring out how to set things up and you'd need to read the docs and code to get oriented :) https://m.soundcloud.com/user-726556259/sherlock-wavegrad-sa... Links to the models are here: https://discourse.mozilla.org/t/creating-a-github-page-for-…

This is actually quite impressive too, significantly better than the last time I looked into Mozilla TTS. Roughly how much audio does "two novels" equate to?

Re: Deep-learning text-to-speech tool for generating voices of various characters

#86

Earlier quoted context omitted.

There's Mozilla TTS https://github.com/mozilla/TTS Here's a sample from a TTS model + vocoder I released for it. I've no wish to deter the motivated, but it'd take a bit of figuring out how to set things up and you'd need to read the docs and code to get oriented :) https://m.soundcloud.com/user-726556259/sherlock-wavegrad-sa... Links to the models are here: https://discourse.mozilla.org/t/creating-a-github-page-for-…

This is actually quite impressive too, significantly better than the last time I looked into Mozilla TTS. Roughly how much audio does "two novels" equate to?

It's about 32 hours of audio.

As some of the audio is read in different accents to the main accent used, ideally the different accent audio would have been removed. Doing so would be expected to help with voice quality, reducing the overall amount used and, as a bonus, cutting training time too.

Re: Deep-learning text-to-speech tool for generating voices of various characters

#87

Earlier quoted context omitted.

There's Mozilla TTS https://github.com/mozilla/TTS Here's a sample from a TTS model + vocoder I released for it. I've no wish to deter the motivated, but it'd take a bit of figuring out how to set things up and you'd need to read the docs and code to get oriented :) https://m.soundcloud.com/user-726556259/sherlock-wavegrad-sa... Links to the models are here: https://discourse.mozilla.org/t/creating-a-github-page-for-…

This is actually quite impressive too, significantly better than the last time I looked into Mozilla TTS. Roughly how much audio does "two novels" equate to?

Here's another sample with the same model+vocoder, this time reading from a Wikipedia article: https://m.soundcloud.com/user-726556259/q-learning-wavegrad-...

Re: Deep-learning text-to-speech tool for generating voices of various characters

#88
post #18
post #17

Earlier quoted context omitted.

You just sort of assume that this is correct? The person[1] running this comes across as a severely unstable character, that number is probably hyperbole. [1] https://twitter.com/fifteenai

Not a hyperbole – I can provide proof if you'd like.

Would you be willing to explain how you can justify offering this for free? I’ve subbed to the patreon, but that’s less than a drop in the bucket compared to the ~$10k you say this month will cost.
Post reply on HN