I was very impressed with the TTS examples in the original DeepMind article ( https://deepmind.com/blog/wavenet-generative-model-raw-audio... ). Can someone elaborate on the usefulness of this implementation for Text-to-Speech? I'm keen to experiment with voice synthesis. I want to create dialog, from multiple voice sources, for some characters in a VR application that I'm working on. Perhaps this lib is a better opt…
Here's a good TTS system: