Sound Event Detection with Adaptive Frequency Selection

doi:10.48550/arXiv.2105.07596

Sound Event Detection with Adaptive Frequency Selection

In this work, we present HIDACT, a novel network architecture for adaptive computation for efficiently recognizing acoustic events. We evaluate the model on a sound event detection task where we train it to adaptively process frequency bands. The model learns to adapt to the input without requesting all frequency sub-bands provided. It can make confident predictions within fewer processing steps, hence reducing the amount of computation. Experimental results show that HIDACT has comparable performance to baseline models with more parameters and higher computational complexity. Furthermore, the model can adjust the amount of computation based on the data and computational budget.

Publication:

arXiv e-prints

Pub Date:

May 2021

DOI:

10.48550/arXiv.2105.07596

arXiv:

arXiv:2105.07596

Bibcode:

2021arXiv210507596W

Keywords:

Computer Science - Sound;
Electrical Engineering and Systems Science - Audio and Speech Processing

E-Print:

Accepted by IEEE Workshop on Applications of Signal Processing to Audio and Acoustics 2021

NASA/ADS

Sound Event Detection with Adaptive Frequency Selection

Abstract