https://research.google.com/audioset/
There's a huge amount to discuss in the audio domain... But for a starting place, using ResNet on spectrograms to build a binary classifier is a good place to start.
11–20 of 30 posts
https://research.google.com/audioset/
There's a huge amount to discuss in the audio domain... But for a starting place, using ResNet on spectrograms to build a binary classifier is a good place to start.
Just found this thread on the fast.ai forum yesterday that may help: https://forums.fast.ai/t/deep-learning-with-audio-thread/381...
aubio and librosa are two excellent MIR (music information retrieval) tools I can recommend from personal use. They can both be implemented for real-time audio using pyaudio or similar. https://aubio.org/doc/latest/ https://librosa.github.io/librosa/
https://www.kdnuggets.com/2016/09/urban-sound-classification...
I don't know if this is off topic but would it be possible to remove the sound of mechanical keyboards with ML in realtime from a VOIP stream? Sell the technology to Discord and profit.
I don't know if this is off topic but would it be possible to remove the sound of mechanical keyboards with ML in realtime from a VOIP stream? Sell the technology to Discord and profit.
Is this a big problem? I thought people loved mechanical keyboard sounds.