Live data from Hacker News

Voice Synthesis for in-the-Wild Speakers via a Phonological Loop

ytaigman.github.io

1–10 of 27 posts

Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop

#8

Similar in quality to Lyrebird https://soundcloud.com/user-535691776/dialog Google WaveNet sounds almost perfect in comparison: https://deepmind.com/blog/wavenet-generative-model-raw-audio...

Some of the generated speech clips are unsettlingly robotic while WaveNet sounds passable, but it's the piano compositions that I found unnerving. I can't explain how but randomly generated music sounds so hollow and cold.

Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop

#10
post #8

Similar in quality to Lyrebird https://soundcloud.com/user-535691776/dialog Google WaveNet sounds almost perfect in comparison: https://deepmind.com/blog/wavenet-generative-model-raw-audio...

Some of the generated speech clips are unsettlingly robotic while WaveNet sounds passable, but it's the piano compositions that I found unnerving. I can't explain how but randomly generated music sounds so hollow and cold.

Note that WaveNet was not trained "in the wild" (like, on celebs) but rather on a speech style dataset
Post reply on HN