arXiv · 2105.09279
Unsupervised Discriminative Learning of Sounds for Audio Event Classification
Abstract
Recent progress in network-based audio event classification has shown the benefit of pre-training models on visual data such as ImageNet. While this process allows knowledge transfer across different domains, training a model on large-scale visual datasets is time consuming. On several audio event classification benchmarks, we show a fast and effective alternative that pre-trains the model unsupervised, only on audio data and yet delivers on-par performance with ImageNet pre-training. Furthermore, we show that our discriminative audio learning can be used to transfer knowledge across audio datasets and optionally include ImageNet pre-training.
Explore related subjects
Keep this discovery
Sascha Hornauer, Ke Li, Stella X. Yu, Shabnam Ghaffarzadegan, Liu Ren. 2021-05-19. Unsupervised Discriminative Learning of Sounds for Audio Event Classification. https://doi.org/10.1109/icassp39728.2021.9413482
Cite the original work for its findings. Save a collection to share your selection of sources.