arXiv · 2610.03381
Multiclass Speech Classification Under Noise Disparity
Abstract
We investigate noise disparity in multi-class speech classification tasks and develop a strategy to prevent classifiers from exploiting class-dependent noise characteristics. Building on our previous work for binary classification, we introduce a multi-class cross-augmentation scheme that exposes each class to the noise characteristics of the other classes, thereby removing the association between individual noise conditions and class labels. We compare this training-based approach with speech enhancement (SE) as a preprocessing strategy, which aims to suppress noise-related cues directly from the input. Experiments on multi-class emotion recognition show that cross-augmentation effectively mitigates the effect of noise disparity across a range of signal-to-noise-ratios, while SE has a detrimental effect on model performance.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Mahdi Amiri, Sayantan Biswas, Mingchi Hou, Pascal Frossard, Ina Kodrasi. 2026-10-02. Multiclass Speech Classification Under Noise Disparity. https://arxiv.org/abs/2610.03381
Cite the original work for its findings. Save a collection to share your selection of sources.