Live data from Hacker News

Deep-learning text-to-speech tool for generating voices of various characters

15.ai

1–10 of 88 posts

Re: Deep-learning text-to-speech tool for generating voices of various characters

#4

will this be open source eventually?

https://twitter.com/fifteenai/status/1342304487474606081

found an answer.

"There's no point in releasing a poorly done model, and to do so for the sake of popularity would be despicable. My goal is to achieve indistinguishability, which I certainly know is possible. Anything short of near-perfection is unacceptable. "

Re: Deep-learning text-to-speech tool for generating voices of various characters

#6
From the about section:

> How much does maintaining the servers cost? > It depends on the amount of traffic, but the minimum baseline is around several thousands of US dollars every month. This is expected as inference is very GPU intensive and a sufficient number of instances need to be spun up to handle thousands of requests coming in every minute. Everything is paid out of pocket.

Wow, impressive commitment for something that's free.

Re: Deep-learning text-to-speech tool for generating voices of various characters

#10

Welp, after messing around with a few voices I was completely impressed with Glados's. This is really cool because I have no idea how the character's voice was synthesized, but apparently ML can do it for me so props to that.

I'm pretty sure the real Glados voice effect is mostly pitch correction and formant shifting. You can do it with Melodyne at least (which, to be fair, is also computer magic-- just a different kind than this one!)

I just found a video on YT with an example of recreating this in Melodyne: https://youtu.be/1oQn66gvwKA

Post reply on HN