If I shutdown every voice other than the optimist's one in my head, this, along with other recent AI research, will mark the advent of never-seen-before role play game possibilities. If the current pace of progress continues, we'll see games with complete narrative freedom for players, where you aren't limited to pre-written answers anymore, but can actually talk to in-game characters with your actual voice, goals, a…
I’m an amateur music producer and vocals are by far the toughest part of making music. I have to find a singer, convince them to work with me (I am an amateur and not particularly good tbh), and book studio space because its very tough to get a clean recording at home. I’m hoping that like digital instruments, I’ll be able to splice in digital voices instead of finding singers.
Audiobox: Meta's new foundation research model for audio generation
31–40 of 90 posts
Re: Audiobox: Meta's new foundation research model for audio generation
#32Re: Audiobox: Meta's new foundation research model for audio generation
#33Earlier quoted context omitted.
I’m an amateur music producer and vocals are by far the toughest part of making music. I have to find a singer, convince them to work with me (I am an amateur and not particularly good tbh), and book studio space because its very tough to get a clean recording at home. I’m hoping that like digital instruments, I’ll be able to splice in digital voices instead of finding singers.
This already exists, ex. Audimee: https://audimee.com/
RVC is so easy anyone can spin up a website for it. No moat. Over a hundred thousand trained weights files in the open, so it's easy to bootstrap.
Re: Audiobox: Meta's new foundation research model for audio generation
#34Does anyone have suggestions for how to integrate this into your tech stack via an internal API? Interested to hear the varying thoughts on this. From what I softly understand is that the model weights have to be swapped or altered per se to be able to commercially reuse this. Correct me if I'm wrong.
Re: Audiobox: Meta's new foundation research model for audio generation
#35If I shutdown every voice other than the optimist's one in my head, this, along with other recent AI research, will mark the advent of never-seen-before role play game possibilities. If the current pace of progress continues, we'll see games with complete narrative freedom for players, where you aren't limited to pre-written answers anymore, but can actually talk to in-game characters with your actual voice, goals, a…
Re: Audiobox: Meta's new foundation research model for audio generation
#36Earlier quoted context omitted.
This already exists, ex. Audimee: https://audimee.com/
And the millions of other RVC websites. Musicfy, Uberduck, Coversai, Kitsai, FakeYou, Voicemyai, Voicify, Bangerapp, Tryreplay, Weightsgg ... RVC is so easy anyone can spin up a website for it. No moat. Over a hundred thousand trained weights files in the open, so it's easy to bootstrap.
"The RVC model is a Retrieval-based Voice Conversion system using AI for high-quality voice cloning. It utilizes artificial intelligence to modify or clone voices in real-time." Source: https://speechify.com/blog/rvc-vocal-models/
Re: Audiobox: Meta's new foundation research model for audio generation
#37Support open source models by celebrating their release and pressuring companies to release them, and oppose closed source AI or face a very bleak future for you and your descendants.
You may be having fun with “Open” AI’s API today, but you’re supporting and celebrating the collapse of society into megacap AI elites and a majority paying for metered access to old technology.
Re: Audiobox: Meta's new foundation research model for audio generation
#38Re: Audiobox: Meta's new foundation research model for audio generation
#39If I shutdown every voice other than the optimist's one in my head, this, along with other recent AI research, will mark the advent of never-seen-before role play game possibilities. If the current pace of progress continues, we'll see games with complete narrative freedom for players, where you aren't limited to pre-written answers anymore, but can actually talk to in-game characters with your actual voice, goals, a…
Re: Audiobox: Meta's new foundation research model for audio generation
#40VR is gonna get wild in like 5 years if they keep this up
There will be people that will spend almost every waking hour with one of those things attached to their face if they can also make this device lightweight and comfortable