Can someone reccomend to me: a service that will generate a loopable engine drone for a "WWII Plane Japan Kawasaki Ki-61"? It doesn't have to be perfect, just convincing in a hollywood blockbuster context, and not just a warmed over clone of a Merlin engine sound. Turns out Suno will make whatever background music I need, but I want a "unique sound effect on demand" service. I'm not convinced voice AI stuff is sustai…
Audio is the one area small labs are winning
41–50 of 109 posts
Re: Audio is the one area small labs are winning
#42Earlier quoted context omitted.
I refuse to believe that none of these people ever heard of Nyquist, and that noone was able to come up with "ayyy lmao let's put a low pass on this before downsampling". Edit: 2 day old account posting stuff that doesn't pass the sniff test. Hmmmm... baited by a bot?
in ~30 years of my work in DSP domain, I've seen insane amount of ways to do signal processing wrong even for simplest things like passing a buffer and doing resampling. The last example I've seen in one large company, done by a developer lacking audio/DSP experience: they used ffmpeg's resampling lib, but, after every 10ms audio frame processed by resampler, they'd invoke flush(), just for the sake of convenience of…
Re: Audio is the one area small labs are winning
#43[flagged]
[flagged]
The reason it matters is that soon, any time somebody sees a comment they don't like or think is stupid, they'll just say, "eh a bot said that," and totally dilute the rest of the discussion, even if the comment was real.
Re: Audio is the one area small labs are winning
#44Re: Audio is the one area small labs are winning
#45They’ll wait for progress to be made and then buy the capability/expertise/talent when the time is right.
Re: Audio is the one area small labs are winning
#46[flagged]
Also, while the author complains that there is not a lot of high quality data around [0], you do not need a lot of data to train small models. Depending on the problem you are trying to solve, you can do a lot with single-digit gigabytes of audio data. See, e.g., https://jmvalin.ca/demo/rnnoise/ [0] Which I do agree with, particularly if you need it to be higher quality or labeled in a particular way: the Fisher data…
If one asks ~~nice~~ expensive enough they can even get isolated multitracks or teleprompter feeds together with the audiovisual tracks. Heck, if they wanted they could set up dedicated transcription teams for the plethora of podcasts with the costs somewhere in the rounding error range. But you can't siphon that off of torrents and paying for training material goes against the core ethics of the big players.
Too bad you can't really scrape tiktok/instagram reels with subtitles... Oh no, oh no, oh no no no no
Re: Audio is the one area small labs are winning
#47Re: Audio is the one area small labs are winning
#48Earlier quoted context omitted.
[flagged]
What's the point of saying that without backing it up? Either you think it's so obvious it doesn't need backing up (in which case you don't need to say it), or ...? The reason it matters is that soon, any time somebody sees a comment they don't like or think is stupid, they'll just say, "eh a bot said that," and totally dilute the rest of the discussion, even if the comment was real.
Re: Audio is the one area small labs are winning
#49Re: Audio is the one area small labs are winning
#50https://www.daily.co/blog/benchmarking-stt-for-voice-agents/