/edit On a more serious node. I already see the 24/7 lofi girl streaming generated music. The sample[1] on lofi sounds pretty good.
[1]https://dl.fbaipublicfiles.com/audiocraft/webpage/public/ass... "Lofi slow bpm electro chill with organic samples"
21–30 of 335 posts
/edit On a more serious node. I already see the 24/7 lofi girl streaming generated music. The sample[1] on lofi sounds pretty good.
[1]https://dl.fbaipublicfiles.com/audiocraft/webpage/public/ass... "Lofi slow bpm electro chill with organic samples"
> MusicGen, which was trained with Meta-owned and specifically licensed music, generates music from text-based user inputs, while AudioGen, which was trained on public sound effects, generates audio from text-based user inputs. Meta is really clearly trying to differentiate themselves from OpenAI here. Open source + driving home "we don't use data we haven't paid for / don't own".
Bully "Open"AI into rebranding.
The demos are great. Could someone explain what’s in it for Meta open sourcing all these models?
Earlier quoted context omitted.
Commoditize Your Complement? https://gwern.net/complement
What is it a complement to though?
Wonder how far off the whole "generate music based on your existing music library" thing is going to be? That'll make musicians happy with big tech as well, just like artists are. *sigh*
The Record labels are far , far more litigious than the art community.
I suppose they might try, anyway.
Curious if I’m alone in that.
(At the bottom https://audiocraft.metademolab.com/musicgen.html)
For what it’s worth though, the voice based examples sound dramatically better with MBD
> MusicGen, which was trained with Meta-owned and specifically licensed music, generates music from text-based user inputs, while AudioGen, which was trained on public sound effects, generates audio from text-based user inputs. Meta is really clearly trying to differentiate themselves from OpenAI here. Open source + driving home "we don't use data we haven't paid for / don't own".
Furthermore, MusicGen's weights are licensed CC-BY-NC, which is effectively a nonlicense as there is no noncommercial use you could make of an art generator[1]. This is not only a 'weights-available' license, but it's significantly more restrictive than the morality clause bearing OpenRAIL license that Stability likes to use[2].
[0] https://github.com/facebookresearch/llama/blob/main/MODEL_CA...
[1] https://github.com/facebookresearch/audiocraft/blob/main/LIC...
[2] These are also very much Not Open Source™ but the morality clauses in OpenRAIL are at least non-onerous enough to collaborate over.
The demos are great. Could someone explain what’s in it for Meta open sourcing all these models?
While the datasets used for training AudioGen aren't available, is there any kind of list where one can review the tags or descriptions of the sounds on which the model was trained? Otherwise how do you know what kinds of sounds you can reasonably expect AudioGen to be capable of generating? And what happens if you request a sound which is too obscure or something not found in the dataset?
What are AudioGen's capabilities regarding spatial positioning? First example: can it generate a siren that starts in front and moves left to right and complete a full circle around the listener? Second example: can it do the same siren but on the Y axis, so it start at the front, it goes over the listener and then it goes under them to complete the circle?