Live data from Hacker News

Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

stability.ai

61–70 of 210 posts

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#61

Earlier quoted context omitted.

This is why it’s vital that AI is openly available. Imagine a world where Spotify is the only company that can do that, and they use it to make sure they never pay royalties again.

How is Spotify for finding new music based on your tastes? I’ve only used Amazon and Pandora; Amazon is quite poor, Pandora is pretty good. I suspect (although, without proof) that if a service can’t suggest new music, it will have trouble generating new music as well. Anyway, I very much would rather run this sort of thing locally. You could just manually set your taste profile. Plus, music can be quite personal, im…

Similarly to others, I haven't really used other things in a LONG time but Spotify's Discover Weekly playlist is usually a list of bops that I enjoy a lot. I frequently end up adding a huge portion of them to my liked songs and to regularly used playlists!

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#62
post #36

Earlier quoted context omitted.

How is Spotify for finding new music based on your tastes? I’ve only used Amazon and Pandora; Amazon is quite poor, Pandora is pretty good. I suspect (although, without proof) that if a service can’t suggest new music, it will have trouble generating new music as well. Anyway, I very much would rather run this sort of thing locally. You could just manually set your taste profile. Plus, music can be quite personal, im…

I find the discover weekly playlists that are made for me are pretty hit or miss, overall I have found many news songs I like with their help.

Same! New songs, new artists, new genres, it's pretty cool! I agree it can be hit or miss but it feels like the longer I use it and the older, and more mature the platform gets, the more hit discover weekly ends up being!

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#63
post #49
post #32

Earlier quoted context omitted.

Stability is great but Meta's MusicGen is available with code and weights while this isn't so that's a really odd place to make that comparison and complaint.

Before stable diffusion, nobody released weights at all. Meta et al only started sharing their models with the world when they realized how fast a developer ecosystem was building around the best models. Without stability, all of AI would still be closed and opaque.

GPT-2, GPT-J, XLNET, BERT, Longformers and T5 were all freely available before Stable Diffusion was even a press release.

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#64
post #36

Earlier quoted context omitted.

How is Spotify for finding new music based on your tastes? I’ve only used Amazon and Pandora; Amazon is quite poor, Pandora is pretty good. I suspect (although, without proof) that if a service can’t suggest new music, it will have trouble generating new music as well. Anyway, I very much would rather run this sort of thing locally. You could just manually set your taste profile. Plus, music can be quite personal, im…

I find the discover weekly playlists that are made for me are pretty hit or miss, overall I have found many news songs I like with their help.

Similar - I found many good new things. I like the discovery playlists, even with the tracks I don't like. If they never missed, how could they ever suggest something actually different and exciting? It's "discover" not "average of what you already like".

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#65

Does this model support / "understand" concepts of spatial audio? For example, something like "an alarm moving around you in a circle". When AudioGen was announced this was my first question, but from what I've been able to test the model just ignores spatial audio prompts. Unfortunately I haven't been able to find any discussion or interest in online discussion about the importance / significance of spatial audio. W…

Dolby wouldn't appreciate it.

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#66

Now imagine Spotify using this to generate individual earworms for everybody based on their personal tastes (likes, playlists). Yes, AI is partly hype, but had someone told me this even two years ago, I wouldn't have believed it.

Spotify does not need to generate a tailored earworm for me. It could already suggest songs that I like based on my personal taste out of their 100-million-songs catalog - and it's absolutely unable to do it.

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#67

The solo piano was interesting because of how clean it is. I can imagine going from that sample to a score without too much difficulty. Once it's in a symbolic format it becomes much more flexible and re-usable. While this does not seem to be the trend I hope more gen ai in the audio and visual realms start to produce more structured / symbolic output. For example, if I were Adobe I would be training models, not to o…

That raises an interesting difference between cleaning AI-generated sound and cleaning ordinary recordings. In an ordinary recording, there is an objective reality to discover -- a certain collection of voices was summed to create a signal. With (most? the best?) existing AI audio generation, the waveform is created from whole cloth, and extracting voices from it is an act of creation, not just discovery.

I've come across AI-generated music that outputs something like MIDI and controls synthesizers. Its audio quality was crystal-clear, but the music was boring. That's not to say the approach is a dead-end, of course -- and indeed, as a musician, the idea of that kind of output is exciting. But getting good data to train something that outputs separate MIDI-ish voices seems much harder than getting raw audio signals.

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#68

It's interesting tech but none of the musical pieces impressed me (I play multiple instruments and have written and arranged music), most sounded too repetitive and not very imaginative. This is also an issue with diffusion based art AI in general, its good at a limited set of things but gets rather repetitive after a while. I could see using this as background music where quality is not important, like in games, tho…

Yes, while working on my AI Melodies Assistant project, it quickly became clear that generating a pleasant but boring music isn't too difficult. To create a catchy tune, an element of surprise is essential. In the end, I was able to use it as an assistant to compose 60 melodies that I'm happy with (https://www.melodies.ai/)

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#70

Now imagine Spotify using this to generate individual earworms for everybody based on their personal tastes (likes, playlists). Yes, AI is partly hype, but had someone told me this even two years ago, I wouldn't have believed it.

Spotify does not need to generate a tailored earworm for me. It could already suggest songs that I like based on my personal taste out of their 100-million-songs catalog - and it's absolutely unable to do it.

Building a tailored earworm might actually be easier.
Post reply on HN