arXiv · 2506.18691
Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders
Abstract
Multichannel speech enhancement algorithms are essential for improving the intelligibility of speech signals in noisy environments. These algorithms are usually evaluated at the utterance level, but this approach overlooks the disparities in acoustic characteristics that are observed in different phoneme categories and between male and female speakers. In this paper, we investigate the impact of gender and phonetic content on speech enhancement algorithms. We motivate this approach by outlining phoneme- and gender-specific spectral features. Our experiments reveal that while utterance-level differences between genders are minimal, significant variations emerge at the phoneme level. Results show that the tested algorithms better reduce interference with fewer artifacts on female speech, particularly in plosives, fricatives, and vowels. Additionally, they demonstrate greater performance for female speech in terms of perceptual and speech recognition metrics.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Nasser-Eddine Monir, Paul Magron, Romain Serizel. 2025-06-23. Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders. https://arxiv.org/abs/2506.18691
Cite the original work for its findings. Save a collection to share your selection of sources.