Live data from Hacker News

MusicGen: Simple and controllable music generation

ai.honu.io

291–300 of 340 posts

Re: MusicGen: Simple and controllable music generation

#291

I mean this is shockingly good. The longer “lofi” example at the bottom sounds like it could have been a Boards of Canada demo. I’m very impressed.

On sound quality, yes. But neither this nor any of the others offered for comparison feels creative in the slightest. By contrast OpenAI's JukeBox, which is maybe two years old now, really comes up with fascinating ideas.

No, not just on sound quality. If that was an outro for a Boards of Canada or Ulrich Schnauss demo, I would not have batted an eye.

Re: MusicGen: Simple and controllable music generation

#292

Earlier quoted context omitted.

how is this fundamentally different than what humans do?

because its a computer not a human

perhaps you had difficulty understanding the question -- how is what the computer is doing, different from what a human does?

Re: MusicGen: Simple and controllable music generation

#293
post #205

(This is not meant to be an anti-ai-generated-art rant. It's coming whether we like it or not. But some of the motives in this thread confuse me.) Music producer here with an honest question to those saying "this will provide me with a simple soundtrack/background music for $PROJECT" Have any of you checked out / made offers on music production subreddits? Or other music subreddits? various music production discords?…

I want to do this but I'm scared of the backlash of "you're being exploitative!!!!".

I know the people who say that mean well, but it totally overlooks both how much the culture does (as you say) want to provide their art for projects to use it and create value together, and... reality. Shouting at everyone isn't the way to get them onboard, but shout they do, and it's one country in particular that seems to scream the most.

I'm in the UK and I can't walk down the street without tripping over producers, so maybe the way around the angsty people is finding them in person? Or... we just use AI. The robots solving our social issues is probably a thing.

Re: MusicGen: Simple and controllable music generation

#294
post #130

Earlier quoted context omitted.

Humans creating and sharing new technologies (and ideas, and works of art, etc.) across societies _is_ a force of nature.

“Force of nature” generally means some phenomenon of physics or some natural disaster outside of human control, which is what I meant. “A big and powerful cool thing” is not what I meant, and not what force of nature usually means.

Life is a force of nature. Evolution is a force of nature. The development of society (across species, not just human) is a force of nature. The development of technology is an aspect of that. In other words, macro-economic trends such as the development of automation ARE in fact an aspect of evolution, aka, a force of nature.

Doesn't matter if that's not the sense in which the phrase is used, these things are arising out of the collective unconscious, not as the result of mere individual will.

Re: MusicGen: Simple and controllable music generation

#295
post #195
post #49

So, CC-BY-NC licensed model weights, and they've made sure to license the training data. And some jurisdictions are saying that copyright cannot be claimed on the output of such models. Oh, to be a fly on the wall in RIAA corporate offices… Sans schadenfreude, I think this (depending on inference speed) could be perfect for dynamic content in games (including IRL games: LARP, escape rooms, table top games, etc.)

>Oh, to be a fly on the wall in RIAA corporate offices… In all likelihood, they're ok with events. Games were never anywhere near their main revenue stream. Now the labour costs on what they're actually selling are dropping to zero. RIAA's future: 1) Use AI to fake a band. 2) Use AI to write music (maybe even lyrics). Don't really care if the AI is any good. 3) Distribute output widely, note that copyright still appl…

Who plays the live shows?

Re: MusicGen: Simple and controllable music generation

#296
post #130

Earlier quoted context omitted.

“Force of nature” generally means some phenomenon of physics or some natural disaster outside of human control, which is what I meant. “A big and powerful cool thing” is not what I meant, and not what force of nature usually means.

Life is a force of nature. Evolution is a force of nature. The development of society (across species, not just human) is a force of nature. The development of technology is an aspect of that. In other words, macro-economic trends such as the development of automation ARE in fact an aspect of evolution, aka, a force of nature. Doesn't matter if that's not the sense in which the phrase is used, these things are arisin…

You just made that up. That's not what the phrase "force of nature" means.

Re: MusicGen: Simple and controllable music generation

#297

Do any of these generative music ML framework export as midi instead of .wav or .mp3? That would be 1000x times more useful as the quality we're reaching is good enough. I can imagine an VSTi that just takes a prompt and generates midi tracks. Something like this is surely coming in the next couple of years.

That was the way almost all ML research on music was done until recently- train for MIDi generation with input of other MIDI, or, at an even more fundamental level, just notes. If you go back, tons of ML music papers were written on generating believable sequences of Bach, Mozart, etc. because it was just note prediction.

Since the advent of transformers, and this idea of using text models mapping the natural language space to tagged music samples, and the music tokenizer acting directly on sampled audio stream (the bits of a .wav file, essentially) all the cutting edge work is going that route. Because it is producing high quality, finished audio streams directly. And I think part of it is because there is way, way more training data for actual audio than there is for MIDI alone (there are tons of free midi sites out there but a lot of it is garbage and it pales in comparison to what is already sampled and tagged in real audio libraries).

I imagine what will happen is... within two to three years these LM-transformer-music models will get so good that the audio will be damn near spotless sounding, and there will be additional methods developed to synthesize with more control directly with the models, to the point where wanting MIDI so you can use your own HQ synth isn't needed, because if you want "the lead synth to sound less digital and more like a classic Minimoog Model D" you just add that to another 'Music2Music" pass and out pops your sound.

For those who still want MIDI there is still work being done on traditional audio-to-MIDI modeling and I think you'd wind up just using that in the chain.

Re: MusicGen: Simple and controllable music generation

#298
post #195

Earlier quoted context omitted.

>Oh, to be a fly on the wall in RIAA corporate offices… In all likelihood, they're ok with events. Games were never anywhere near their main revenue stream. Now the labour costs on what they're actually selling are dropping to zero. RIAA's future: 1) Use AI to fake a band. 2) Use AI to write music (maybe even lyrics). Don't really care if the AI is any good. 3) Distribute output widely, note that copyright still appl…

Who plays the live shows?

That's what actors + pre-recordings are for (we could try using AI to generate the music live but that would make the actors' work more complicated and add more fail modes, why bother?). Note that using recordings already happened in the pre-AI era:

https://blabbermouth.net/news/anvils-lips-on-bands-using-bac...

Re: MusicGen: Simple and controllable music generation

#300

This is incredible!! For all the “AI is stealing our music” naysayers here consider that all art is derivative or it lacks context and makes it nonsensical, and artists learn too.

I don't know why you're mentioning art being derivative, but that's not the thing that worries me.

What worries me, is that a good enough model will take away the incentive to write music for many, and as a consequence it will also remove performers. This will reduce demand on music teaching and instruments, which will then both become nearly inaccessible. Since learning music isn't a question of following a few youtube videos, this will leave the world with just AI music.

Jazz and classical music are probably exempt, since it relies on subsidies, their audiences care about the actual performance, and AI compositions will not draw enough of a crowd to make it financially interesting.

But popular music will suffer, and that's what makes development of these models straight evil.

Post reply on HN