Live data from Hacker News

How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

zhinit.dev

51–60 of 93 posts

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#51
post #35

Earlier quoted context omitted.

For me creating the exact sound is not very interesting from sound designing perspective. You can always sample the real instrument. Like physical modeling synthesis, the interesting part is to compress the sound to some parameters that you can tweak and generate new sounds Another approach is VAE, which also you give your some latent embedding, you can tweak the embedding to generate new sound. However the meaning o…

"You can always sample the real instrument." This doesn't really work on instruments like guitars. Open D sounds way different than fretted D on the E string. Timbre changes with position and it's one of the ways I determine where a player's hands are on the neck when I'm trying to play their song.

That is not something inherent in guitars themselves, it is the norm in steel string guitars and the fan-braced/Spanish guitar but mostly because that is the norm for all those mass produced guitars which make up the bulk of guitars. On steel string you can often greatly decrease this quality just by switching to flatwounds, this is part of the flatwound sound, it shifts the timbrel content into the players technique but if you want much timbrel content with flatwounds you need heavy strings and a high action, and the hand strength and technique that sort of setup requires.

Before the rise of the steel string and the Spanish guitar, guitars tended to be more even across their range and also had less bass which helped even them out, and now that sound is what we are used to. There have always been niches that wanted that more even sound, but for most that just makes it more difficult to play all that music that developed around these quirks, so they remain niches.

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#53

Earlier quoted context omitted.

I'm not doing fancy AI stuff but I have worked a lot with my own bespoke supercollider system where I record whole fretboards of guitars and then play alternative notes based off of certain rules. For whatever dumb reason though, the most natural sounding thing is really just playing, e.g., any random D4 from its possibilities at any given moment. Timbral differences also exist depending on force, the manner plucked,…

No, I'm going off the timbral differences - same way I identify which pickup position is being used. There's a specific 'thickness' I cue in on to determine pickup and specific note placement.

Huh, got it. That's pretty cool!

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#54

Slightly off-topic. Now that 1920s jazz music is falling into public domain, has anyone tried to reinvigorate the music using AI and generative adversarial approaches? Pre-1940s music didn't have high-fidelity sound, so the strong bass lines weren't captured. In theory, we could "downgrade" modern recordings to sound like 1920s recordings, then use adversarial techniques to train the machine on how to restore the ant…

It might be easier than that. Are the bass lines totally missing or are they just very weak? If you can capture a recording using vintage equipment and the placing of it, you can get the system response. Run the original recordings through an inversion of the response and you should get really close. Another possible method is to find the transform between an identical modern recording of the song and use the difference between the two recordings to make your transform.

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#55
post #34
post #10

Earlier quoted context omitted.

For sure! I just added a couple

Did you save any of the "failed" results? I'd love to hear what kind of weird sounds it makes out of distribution (e.g. on the keywords it didn't have much data for).

I just added one with the "techno" keyword in the keywords section. Its pretty weird lol

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#56
post #29
post #25

People who are interested in this application should check synplant[0]. It has a ML technology called "Genopatch" which gives you 2 functionality: 1. you can try to describe a sound with some tags and it will try to generate a sound to capture the feeling of these tags 2.you can feed it with a sound sample and it will try to re-synthesize the sound with its synth engine. Though the end result will usually be just a "…

How far are we from getting a general model that can resynthesize any instrumental audio sound without fiddling with any knobs, so that we can recreate instruments we hear from any song? Seems like it should exist by now?

Fiddling with the knobs is the fun part. https://www.youtube.com/watch?v=la2u4VlGwbQ

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#57
post #45
post #37

This has been done years ago. See https://audialab.com/products/emergent-drums-2/ for instance.

Interesting! I had not seen this. On their website they mention diffusion but not the other models so it might not be identical but its definitely similar.

That's fair :) I also could've phrased my comment a bit more politely.

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#58
post #54

Slightly off-topic. Now that 1920s jazz music is falling into public domain, has anyone tried to reinvigorate the music using AI and generative adversarial approaches? Pre-1940s music didn't have high-fidelity sound, so the strong bass lines weren't captured. In theory, we could "downgrade" modern recordings to sound like 1920s recordings, then use adversarial techniques to train the machine on how to restore the ant…

It might be easier than that. Are the bass lines totally missing or are they just very weak? If you can capture a recording using vintage equipment and the placing of it, you can get the system response. Run the original recordings through an inversion of the response and you should get really close. Another possible method is to find the transform between an identical modern recording of the song and use the differe…

> Are the bass lines totally missing or are they just very weak?

I think it might be that it's missing a large part of the lower frequencies, not that entire bass sound is missing. And I guess it'd be hard to faith-fully regenerate those, if we simply don't have a lot of samples.

Re: How to Train a Gen AI Kick Drum Model on Your Old Linux Desktop with 6GB VRAM

#60

Slightly off-topic. Now that 1920s jazz music is falling into public domain, has anyone tried to reinvigorate the music using AI and generative adversarial approaches? Pre-1940s music didn't have high-fidelity sound, so the strong bass lines weren't captured. In theory, we could "downgrade" modern recordings to sound like 1920s recordings, then use adversarial techniques to train the machine on how to restore the ant…

There are plenty of current human bands that play this music really well though...
Post reply on HN