Live data from Hacker News

Lyrebird – An API to copy the voice of anyone

lyrebird.ai

141–150 of 311 posts

Re: Lyrebird – An API to copy the voice of anyone

#141
post #131

Earlier quoted context omitted.

The "Obama" material sounded quite good, but reverb can cover up a multitude of sins...

Huh. That means that recordings of speeches/performances from concert halls are potentially suspicious. Know of any other instances of reverb covering things up? This is interesting!

I haven't worked extensively in voice adaptation, but I learned working in text-to-speech that adding a bit of reverb is quite effective at covering up artifacts.

Something similar seems to be going on in live vocals. If you lack confidence in your own voice, adding a bit of reverb can make it sound much better. Not sure what's going on — whether the reverb jams the critical listening facilities in one's brain or something like that.

Re: Lyrebird – An API to copy the voice of anyone

#142

Cooler : http://www.dtic.upf.edu/~mblaauw/IS2017_NPSS/ https://arxiv.org/abs/1704.03809

Very impressive, but like someone else said there are definite robotic/synthetic moments. I wonder how easy it would be to combine all of the other projects linked in these comments, is there a common interface that could easily combine all of them? Probably not because of commercial concerns...

Re: Lyrebird – An API to copy the voice of anyone

#143
post #29

I wonder how dependent this is on language: can we make Trump speak Chinese using a one minute audio track of him speaking English?

I'd imagine you might get close, but it'd probably work best if you can get the person to use all the phonemes that you want to reproduce. That said depending on how good it is, it might slur other phonemes together to approximate it, which would probably work to give it the accent that the speaker would likely have.

It all sounded sort of slurry and muffled. Maybe if you're imitating a naturally slurred speaker it would be more effective.

Re: Lyrebird – An API to copy the voice of anyone

#145
post #98
post #80

Earlier quoted context omitted.

Forget dead husbands, with this tech, it will be hard to trust anything a politician said. Basically, once they master adding this to video, ANYTHING could be construed against anyone. Want a video of a politician saying "Hitler was right" to cheering masses? Want a video about a president saying it's time to start Nuclear War One? You can make that.

https://www.youtube.com/watch?v=ohmajJTcpNk hf

Hi.

This is your first and only post, you have no submissions or favorites, and your account is 193 days old.

I'm very curious (and perplexed) as to why you have linked a video from elsewhere in this thread with no supporting context regarding its relevance other than "hf".

Re: Lyrebird – An API to copy the voice of anyone

#146

Earlier quoted context omitted.

In the first clip, I'd say 80% of the soundbites were obviously robot-like, but one or two of the "Obama" quotes were startlingly clear - "The good news is, that they will offer the technology to anyone" - I can't hear anything wrong with that in the first clip at all. If they were all that quality I'd say we'd be easily fooled. As a proof of concept this is pretty big.

I can definitely hear issues with that phrase. It has quite robotic drop-offs. Though coming soon: Neural networks to determine whether speech is NN-generated? :P

In a sort of Turing Test where I don't know who's a robot, or where I'm not even expecting a robot, it would probably be a bit harder.

Re: Lyrebird – An API to copy the voice of anyone

#147
post #131

Earlier quoted context omitted.

Huh. That means that recordings of speeches/performances from concert halls are potentially suspicious. Know of any other instances of reverb covering things up? This is interesting!

I haven't worked extensively in voice adaptation, but I learned working in text-to-speech that adding a bit of reverb is quite effective at covering up artifacts. Something similar seems to be going on in live vocals. If you lack confidence in your own voice, adding a bit of reverb can make it sound much better. Not sure what's going on — whether the reverb jams the critical listening facilities in one's brain or som…

Wow. TIL x 2!

Now I understand why I liked adding a bit of reverb when listening to old MOD/IT/S3M audio files - it covered up the "digitalness" of the song structure a bit.

Thanks for the live vocals tidbit too, that's definitely something to file away.

I wonder how far you could push that in a presentational context (ie, when giving speeches), or whether "who left the speakers in 'dramatic cathedral' mode" would happen before "I dunno what they did to the audio but it sounds great". Maybe if the presentation area was fairly open/large it could work; the question is whether it would have a constructive effect.

Re: Lyrebird – An API to copy the voice of anyone

#148

Last week on BBC Radio 4 I heard of a woman who was losing her voice through disease (MND maybe?), a similar system was being anticipated and she was saving voice samples to seed it with. She had been a singer and strongly identified her self with her voice, she wanted to be able to use a speech synthesis system that had her own voice pattern. Apologies if this was already mentioned, but it seems to be a use others h…

If it seemed to you to be a use others hadn't considered, why would you apologize for mentioning it?

Re: Lyrebird – An API to copy the voice of anyone

#150
post #131

Earlier quoted context omitted.

The "Obama" material sounded quite good, but reverb can cover up a multitude of sins...

Huh. That means that recordings of speeches/performances from concert halls are potentially suspicious. Know of any other instances of reverb covering things up? This is interesting!

Also - lowering the bit-rate can coverup other defects (e.g. phone call).

The cadence was a bit off/unnatural, but I'm sure that is not too hard to fix. Phone-in TV/Radio/web shows are about to get very interesting.

Post reply on HN