Too dangerous for PR reasons at least until after November .
If November is the reason, well, this isn't the last November with that concern...
AI speech generator 'reaches human parity' – but it's too dangerous to release
61–68 of 68 posts
Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#62Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#63Ah, the old "it's too dangerous to release" marketing move. Why even tell us about it?
Not everything is a conspiracy. It is very plausible they're speaking genuinely.
Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#64Speech generation has gotten really good, but there's simply no way to faithfully recreate someone's vocal idiosyncracies and cadence with just "a few seconds" of real audio. That's where the models tend to fall short.
Few seconds means less than a minute. That’s not nothing. Look at a clock and talk for a minute — it’s longer than you might think. Do you think you could give a recording of a minute of someone talking to a talented impressionist and they could impersonate that person to some degree? It doesn’t seem that far fetched to me.
Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#65Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#66What is the point of them trying to create this? That something like this would mostly be used to create disinformation and create chaos is easily understood before making something like this. Truly irresponsible
Perfecting the tech for wide-spread use has trade offs; need for caller id, ease of slandering until trust in voice uniqueness recalibrates, all of which is going to change soon anyway but giving only rich/bad actors the tech at first has its own set of trade offs. Head in the sand is the irresponsible way.
Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#67Earlier quoted context omitted.
Try StyleTTS2. You will still have to experiment with the settings a little to get the right level of adherence to the reference speaker’s voice and the emotion content.
Without looking at this, are you sure that this can do speech to speech? Maybe my flaw in searching has been disregarding anything that's called "text to speech" as not also "speech to speech"?
Re: AI speech generator 'reaches human parity' – but it's too dangerous to release
#68Earlier quoted context omitted.
And it's even more plausible this is just a marketing play to build hype. Take, for example, you just made some new super-pathogen in your basement lab. It could kill everyone on the planet. This is obviously pretty dangerous, so do you: A) silently dispose of it and hope nobody else ever makes the mistake of creating it. or B) keep it in the freezer and hold a press release about how you made it but it's too dangero…
No, it's not more plausible.