Lyrebird – An API to copy the voice of anyone
291–300 of 311 posts
Re: Lyrebird – An API to copy the voice of anyone
#292Last week on BBC Radio 4 I heard of a woman who was losing her voice through disease (MND maybe?), a similar system was being anticipated and she was saving voice samples to seed it with. She had been a singer and strongly identified her self with her voice, she wanted to be able to use a speech synthesis system that had her own voice pattern. Apologies if this was already mentioned, but it seems to be a use others h…
If it seemed to you to be a use others hadn't considered, why would you apologize for mentioning it?
Re: Lyrebird – An API to copy the voice of anyone
#293Earlier quoted context omitted.
I'm pretty sure "passing off" requires the seller to be fraudulently claiming the goods are the goods of someone else, that if you up-front say "the voices used are generated by computer algorithm and do not represent any real person" that a claim of passing off would be rendered moot. Trademark/Copyright can't be disclaimed in this way but Passing Off requires active deception AIUI?
How is this not vulnerable to the "Mickey Mouse animated feature film but drawn by 6th graders"?
It might be worth noting Mickey Mouse is a registered trademark and so doesn't need to use the weaker Passing Off law?
Re: Lyrebird – An API to copy the voice of anyone
#294Earlier quoted context omitted.
In my (admittedly limited) experience mixing for both recording and live settings, I can say that the "sounds great" comes a long way before the "dramatic cathedral mode". If you can listen to it and hear reverb (unless you're going for that effect) you're doing it wrong. What you want is a bit of fullness, slightly softer edges at the end of words/sentences. It's similar to the difference between 24fps cinema and 60…
Brains are incredible differential engines not evolved to handle current technology. A standard quality video is just a projection, a high quality video stream on a 4k set running full 122hz is a weird window from wich we don't get stereo depth clues. The brain constantly has to rethink it's not real as we shift the head and the pov doesn't adjust.
Once LCD density allows for VR with 4K (or, if needed, 8K) per eye... yeah :) we'll firmly be in the virtual reality revolution.
Obviously we'll also need tracking and rendering that can keep up but display density is one of the trickier problems right now.
Re: Lyrebird – An API to copy the voice of anyone
#295- legalscreenshot.com
- legalprintscreen.com
I also developed a concept of "Reality Check" similar to Touring Test (when VR and AI becomes so convincing >50% people won't distinguish it from base reality)... Too bad I'm on the corporate network and my personal website is blocked: https://genesis.re/wiki
Aside: do you believe psychedelics should be the part of obligatory astronaut training?
Re: Lyrebird – An API to copy the voice of anyone
#296Earlier quoted context omitted.
The one I listen to will occasionally mispronounce something in a way that a human never would, or say the name of a punctuation mark.
Huh. I see. Do you happen to know what station you use? I'd kind of like to hear this for myself (for the sole reason that I'd like to get an idea of what it sounds like, since that does definitely sound like a TTS).
BTW, I like the Tom voice more than the newer Paul. Paul is more realistic, but is also more soft-spoken and monotonic. Tom has more inflection and sounds more...forceful. I know it's just my imagination, but sometimes Tom sounds annoyed at bad weather :)
Re: Lyrebird – An API to copy the voice of anyone
#297Combined with Face2Face[1] live video impersonation, it is truly time to be very careful verifying videos or even live streams. https://www.youtube.com/watch?v=ohmajJTcpNk
Without a doubt, our concept of personal identity will be completely unreliable within a few generations. Forget about privacy--we will soon have literally no way to verify who we're talking to.
Re: Lyrebird – An API to copy the voice of anyone
#298Earlier quoted context omitted.
Huh. I see. Do you happen to know what station you use? I'd kind of like to hear this for myself (for the sole reason that I'd like to get an idea of what it sounds like, since that does definitely sound like a TTS).
KZZ40,162.45 MHz, Deerfield NH. Note that the stations have several different voices they use for different reports. Now that I think about it, I'm not sure which one I heard the mistakes on - it might have been one of the older ones. BTW, I like the Tom voice more than the newer Paul. Paul is more realistic, but is also more soft-spoken and monotonic. Tom has more inflection and sounds more...forceful. I know it's j…
I see. I can't seem to find an online receiver for that frequency, although I did find that WZ2500 uses or seems to have used that frequency (for Wytheville VA).
I had a look at SDR.hu (a site I may or may not have just dug out of Google for the first time), but unfortunately the RTL-SDR receivers I can find seem to focus entirely on 0-30MHz. There are a couple ~400MHz receivers but nothing for ~160MHz.
(I may have fired up the receiver I found in NH and fiddled with it, puzzled, for 10 minutes before realizing the scale is in kHz, not MHz... yay)
> Note that the stations have several different voices they use for different reports.
Right.
> Now that I think about it, I'm not sure which one I heard the mistakes on - it might have been one of the older ones.
That's entirely possible. (But hopefully not. I kind of want to hear. :P)
> BTW, I like the Tom voice more than the newer Paul. Paul is more realistic, but is also more soft-spoken and monotonic. Tom has more inflection and sounds more...forceful. I know it's just my imagination, but sometimes Tom sounds annoyed at bad weather :)
I just learned about this service, I have to admit (I'm in Australia). It sounds really nice to be able to have a computer continuously read out the weather conditions to you as they change. And I can completely relate to the idea of preferring the voice that sounds unimpressed when the weather's bad :D
Re: Lyrebird – An API to copy the voice of anyone
#299Earlier quoted context omitted.
KZZ40,162.45 MHz, Deerfield NH. Note that the stations have several different voices they use for different reports. Now that I think about it, I'm not sure which one I heard the mistakes on - it might have been one of the older ones. BTW, I like the Tom voice more than the newer Paul. Paul is more realistic, but is also more soft-spoken and monotonic. Tom has more inflection and sounds more...forceful. I know it's j…
> KZZ40,162.45 MHz, Deerfield NH. I see. I can't seem to find an online receiver for that frequency, although I did find that WZ2500 uses or seems to have used that frequency (for Wytheville VA). I had a look at SDR.hu (a site I may or may not have just dug out of Google for the first time), but unfortunately the RTL-SDR receivers I can find seem to focus entirely on 0-30MHz. There are a couple ~400MHz receivers but…
I'm not surprised that they re-use the frequencies. These are local weather stations, only intended to serve a radius of a hundred miles or so (at least here on the east coast). In addition to my local station I can receive the one in Boston, about 50 miles south of me.