Earlier quoted context omitted.
It won't stop it, but will allow enforcement agencies to enforce. Otherwise, they have no legal recourse to do so.
Enforce what? Why should the ability to impersonate a persons voice suddenly become a crime in itself? Should we arrest Jim Carrey? Isn't it when the thing was used to do something else illegal when enforcement is required?
Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
281–290 of 471 posts
Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#282Earlier quoted context omitted.
This is democratising the tech. Otherwise only the intelligence agencies will have it and we will continue to be duped not knowing what is possible.
We don't need to all be able to use the tech for it to be known publicly. Apply your same logic to any other easily misused tech: "We must all have easy access to bio-engineered viruses. Otherwise only..." "We all need to have access to nuclear weapons. Otherwise only..." Not all tech should be in everyone's hands.
That is not possible with or comparable to, things such as bioweapons.
Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#283I recommend you immediately add identity verification (state-issued identification verification), set up appropriate secrets store for PII, and audit trail EVERYTHING your users are doing, storing the contents in a secure location. Yesterday. This service will be used to harm others, shortly. I do think that there are exciting, honest things that can be done with this service but you need to set up some friction for…
> I recommend you immediately add identity verification (state-issued identification verification)
and
> The genie is already out of the bottle. The degree of effort to put this together is low enough that it will be replicated around the world.
are thoughts that end up in the same post?
If the genie is out of the bottle, it’s your proposed solution that everybody that runs a model like this implements bank-style KYC?
What do you propose should happen when this sort of software becomes freely available for everyone? When (not if) that happens, what will your suggestion have accomplished?
Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#284Earlier quoted context omitted.
Cofounder here, What you see in the above demo is a very rate-limited demo of our upcoming model. We realize how dangerous this technology can be and have built a lot of mitigations on our main product (Play.ht) to reduce possible abuse: - We strictly moderate the generated text of any sexual, offensive, racist, or threatening content. It automatically gets detected and blocked. - We built and are offering for free a…
>We strictly moderate the generated text of any sexual, offensive, racist, or threatening content. This won't be the problem. My voice calling my parents asking for money to be sent to a random account will be the problem. And none of that will be sexual, offensive, racist, or threatening. >we are working hard to mitigate that and deploy it safely. How? >we have seen enough genuine use cases What?
Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#285Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#286Earlier quoted context omitted.
Like this example here: https://playground.play.ht/listen/1554 which says: > "Hi Mom, I need some help. Some guys hit me over the head and put me in a van, and they're saying they'll kill me if you don't wire money to this bank account." top class. EDIT this was about one page down on the "see what people are generating" page
On the bright side, it's not a very convincing rendition of a human.
Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#287Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#288Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#289Earlier quoted context omitted.
On the bright side, it's not a very convincing rendition of a human.
I agree. That guy sounds very nonchalant for being in life-threatening distress.
PS on the downvote: sorry if I did hurt someone's feelings, but it's the truth
Re: Launch HN: Play.ht (YC W23) – Generate and clone voices from 20 seconds of audio
#290Earlier quoted context omitted.
The OP mentioned that for so called, "High-fidelity voice cloning", it would take 20 minutes of training. I think a book author would want the best quality possible to reproduce their voice.
Why reproduce their voice? There's no value-add there.
But I'm not sure. Part of why I'd prefer the original author to read a book is that they vocally emphasize certain parts of the book, and I don't think these models could do that at this point.