Deep-learning text-to-speech tool for generating voices of various characters
21–30 of 88 posts
Re: Deep-learning text-to-speech tool for generating voices of various characters
#22will this be open source eventually?
https://twitter.com/fifteenai/status/1342304487474606081 found an answer. "There's no point in releasing a poorly done model, and to do so for the sake of popularity would be despicable. My goal is to achieve indistinguishability, which I certainly know is possible. Anything short of near-perfection is unacceptable. "
Re: Deep-learning text-to-speech tool for generating voices of various characters
#23will this be open source eventually?
https://twitter.com/fifteenai/status/1342304487474606081 found an answer. "There's no point in releasing a poorly done model, and to do so for the sake of popularity would be despicable. My goal is to achieve indistinguishability, which I certainly know is possible. Anything short of near-perfection is unacceptable. "
AI and ML users are massively benefiting from open source but too often refuse to release their data. It's like we're back in the middle ages and alchemy is back in style.
Re: Deep-learning text-to-speech tool for generating voices of various characters
#24Re: Deep-learning text-to-speech tool for generating voices of various characters
#25will this be open source eventually?
https://twitter.com/fifteenai/status/1342304487474606081 found an answer. "There's no point in releasing a poorly done model, and to do so for the sake of popularity would be despicable. My goal is to achieve indistinguishability, which I certainly know is possible. Anything short of near-perfection is unacceptable. "
I do plan to compile and publish my findings in the future, but nothing is set in stone yet. I know that the model can be improved even further, and I'd prefer to be as comprehensive as possible.
Re: Deep-learning text-to-speech tool for generating voices of various characters
#26Earlier quoted context omitted.
https://twitter.com/fifteenai/status/1342304487474606081 found an answer. "There's no point in releasing a poorly done model, and to do so for the sake of popularity would be despicable. My goal is to achieve indistinguishability, which I certainly know is possible. Anything short of near-perfection is unacceptable. "
Megalomania, always a great excuse. AI and ML users are massively benefiting from open source but too often refuse to release their data. It's like we're back in the middle ages and alchemy is back in style.
Re: Deep-learning text-to-speech tool for generating voices of various characters
#27Being able to generate voices for games would enable a lot of interesting indie projects. IMO people should be paying more attention the market implications of products like this than to the social implications. There are a lot of projects that just aren't really feasible right now that could be if this kind of technology was more polished and generally available for commercial/self-hosted use. And in those cases, you don't even need to do inference, makers will likely be willing to mark up their scripts themselves.
Anyway I digress. Congrats, this is really cool!
Re: Deep-learning text-to-speech tool for generating voices of various characters
#28I don't usually expect much from demos like this, but I'm kind of surprised how impressive the results currently are. They're definitely not perfect, you're definitely getting some odd clipping and noise, but this shows a large amount of promise. Being able to generate voices for games would enable a lot of interesting indie projects. IMO people should be paying more attention the market implications of products like…
People will absolutely suffer harm from this tech, but hey, think about the dollars that could be made! No, we should absolutely be paying more attention to the social implications.
Re: Deep-learning text-to-speech tool for generating voices of various characters
#29From the about section: > How much does maintaining the servers cost? > It depends on the amount of traffic, but the minimum baseline is around several thousands of US dollars every month. This is expected as inference is very GPU intensive and a sufficient number of instances need to be spun up to handle thousands of requests coming in every minute. Everything is paid out of pocket. Wow, impressive commitment for so…
You just sort of assume that this is correct? The person[1] running this comes across as a severely unstable character, that number is probably hyperbole. [1] https://twitter.com/fifteenai
I have no reason to disbelieve it.
Re: Deep-learning text-to-speech tool for generating voices of various characters
#30I don't usually expect much from demos like this, but I'm kind of surprised how impressive the results currently are. They're definitely not perfect, you're definitely getting some odd clipping and noise, but this shows a large amount of promise. Being able to generate voices for games would enable a lot of interesting indie projects. IMO people should be paying more attention the market implications of products like…
> people should be paying more attention the market implications of products like this than to the social implications. People will absolutely suffer harm from this tech, but hey, think about the dollars that could be made! No, we should absolutely be paying more attention to the social implications.
I'm not primarily interested about the dollars, I'm interested in allowing communities to do creative things. I think people are looking at this tech like it's only going to be used for deepfakes, and they're underestimating the extent it's going to be used to create voice-acted game mods, animations, anonymization tools, and other creative/helpful projects.
If you're really worried about this stuff though, you can take some comfort in the fact that by far the worst examples on the site are of real-world voices. This is currently technology that as far as I can see is far more suited for generating new voices or voicing cartoon characters with well-defined patterns/inflections than it is for imitating the president.