To me this is very exciting. I'm already working on my own home digital assistant modeled as NeNe Leaks from the Real Housewives to add personality to otherwise boring conversations with a robot. I've been looking at various style transfer techniques, and having something a bit more plug & play will help me focus on the more unique parts. I predict that we'll see more celebrity voices used as conversational interface…
Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
21–27 of 27 posts
Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
#22Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
#23I still think emphasis on a word or syllable is important here as there is far more information than you realize being conveyed with inflection. Consider: I am going to eat the ham sandwich = Me, no one else I am going to eat the ham sandwich = Nothing can stop me I am going to eat the ham sandwich = On my way; got distracted I am going to eat the ham sandwich = In case you doubt my intent I am going to eat the ham s…
I am going to eat the ham sandwich = The sandwich is the reason I am going (to the party, or wherever...)
Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
#24Similar in quality to Lyrebird https://soundcloud.com/user-535691776/dialog Google WaveNet sounds almost perfect in comparison: https://deepmind.com/blog/wavenet-generative-model-raw-audio...
Some of the generated speech clips are unsettlingly robotic while WaveNet sounds passable, but it's the piano compositions that I found unnerving. I can't explain how but randomly generated music sounds so hollow and cold.
Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
#25To me this is very exciting. I'm already working on my own home digital assistant modeled as NeNe Leaks from the Real Housewives to add personality to otherwise boring conversations with a robot. I've been looking at various style transfer techniques, and having something a bit more plug & play will help me focus on the more unique parts. I predict that we'll see more celebrity voices used as conversational interface…
What language are you working in? I've been working in Powershell out of convenience, but am looking to port my speech bot to Node.
Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
#26Earlier quoted context omitted.
Subvocalisation. Your own voice speaking quietly in the background to something else. Ideal for consumer indoctrination.
The voice we hear in our head and the one everyone else hear is starkly different.
Re: Voice Synthesis for in-the-Wild Speakers via a Phonological Loop
#27Earlier quoted context omitted.
Subvocalisation. Your own voice speaking quietly in the background to something else. Ideal for consumer indoctrination.
The voice we hear in our head and the one everyone else hear is starkly different.