arXiv · 2512.03563
State Space Models for Bioacoustics: A Comparative Evaluation with Transformers
Abstract
In this study, we evaluate the efficacy of the Mamba architecture bioacoustics by introducing BioMamba, a Mamba-based audio representation model for wildlife sounds. We pre-train a BioMamba using self-supervised learning on a large audio corpus and evaluate it on the BEANS benchmark across diverse classification and detection tasks. Compared to the state-of-the-art Transformer-based model (AVES), BioMamba achieves comparable performance while significantly reducing VRAM consumption. Our results demonstrate Mamba's potential as a computationally efficient alternative for real-world environmental monitoring.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Chengyu Tang, Sanjeev Baskiyar. 2025-12-03. State Space Models for Bioacoustics: A Comparative Evaluation with Transformers. https://arxiv.org/abs/2512.03563
Cite the original work for its findings. Save a collection to share your selection of sources.