Expressive Speech Synthesis with Tacotron
research.googleblog.com
Expressive Speech Synthesis with Tacotron
1–10 of 11 posts
Re: Expressive Speech Synthesis with Tacotron
#2This means we now live in a world where you can create a recording of Donald Trump saying, "I colluded with the Russians to rig the election," and not only have the voice sound like Trump but also bring along his personal expressive style so that it becomes indistinguishable from Trump himself.
Would love to see these two combined - make an audio-video recording of an actor confessing to election fraud, then use Face2Face to swap in Trump's face and use Tacotron to swap in his voice.
Re: Expressive Speech Synthesis with Tacotron
#3This web demo allows you to enter your own text:
https://cloud.google.com/text-to-speech/
(select US American and Wavenet)
Re: Expressive Speech Synthesis with Tacotron
#4This technology has been around for a year but we only got a few samples. I'm very excited. I use TTS to read back all the text I consume on PC. This web demo allows you to enter your own text: https://cloud.google.com/text-to-speech/ (select US American and Wavenet)
Nope! These two papers are fresh work on prosody modeling. You can see the evolution of work this team has been publishing about here:
https://google.github.io/tacotron/
> This web demo allows you to enter your own text:
That web demo is unrelated to this. It's about a Google Cloud TTS API, which only includes WaveNet, not Tacotron.
https://cloudplatform.googleblog.com/2018/03/introducing-Clo...
Re: Expressive Speech Synthesis with Tacotron
#5Re: Expressive Speech Synthesis with Tacotron
#6https://google.github.io/tacotron/publications/global_style_... Ha!
Re: Expressive Speech Synthesis with Tacotron
#7Re: Expressive Speech Synthesis with Tacotron
#8Re: Expressive Speech Synthesis with Tacotron
#9https://google.github.io/tacotron/publications/global_style_... Ha!