Live data from Hacker News

Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

stability.ai

1–10 of 210 posts

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#3
Thank you for sharing! On a tangent: I'm wondering if there are any good open source models/libraries to reconstruct audio quality. I'm thinking about an end-to-end open source alternative to something like Adobe Podcast [1] to make noisy recordings sound professional. Anecdotally it's supposed to be very good. In a recent search, I haven't found anything convincing. In my naive view this tasks seems much simpler than audio generation and the demand far bigger, since not everyone has a professional audio setup ready at all times.

[1] https://podcast.adobe.com/

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#5
post #3

Thank you for sharing! On a tangent: I'm wondering if there are any good open source models/libraries to reconstruct audio quality. I'm thinking about an end-to-end open source alternative to something like Adobe Podcast [1] to make noisy recordings sound professional. Anecdotally it's supposed to be very good. In a recent search, I haven't found anything convincing. In my naive view this tasks seems much simpler tha…

We've been researching an audio denoiser for music that we will present at the AES conference in October. Description page: https://tape.it/denoising

We'll also publish a webapp where you can use the denoiser for free. Mail me if you want beta access to it (email in profile).

It won't be open-source though, although the paper will of course be public. It will also only reduce noise, and not reconstruct other aspects of audio quality. However, it can do so on any audio (in particular music), not just speech like Adobe Podcast, and it fully preserves the audio quality. It's designed exactly for the use case you want: to make noisy recordings sound professional.

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#8
post #3

Thank you for sharing! On a tangent: I'm wondering if there are any good open source models/libraries to reconstruct audio quality. I'm thinking about an end-to-end open source alternative to something like Adobe Podcast [1] to make noisy recordings sound professional. Anecdotally it's supposed to be very good. In a recent search, I haven't found anything convincing. In my naive view this tasks seems much simpler tha…

We've been researching an audio denoiser for music that we will present at the AES conference in October. Description page: https://tape.it/denoising We'll also publish a webapp where you can use the denoiser for free. Mail me if you want beta access to it (email in profile). It won't be open-source though, although the paper will of course be public. It will also only reduce noise, and not reconstruct other aspects…

denoising seems to fail in the guitar and vocals example

Re: Stable Audio: Fast Timing-Conditioned Latent Audio Diffusion

#10
post #6

The bluegrass one is super weird. I can’t identify exactly why.

You are right it feels off.

The position of the guitar in stereo is all over the place, higher frequency elements appear to come from the left while other parts are more centered.

Post reply on HN