arXiv · 2102.11906
Handling Background Noise in Neural Speech Generation
Abstract
Recent advances in neural-network based generative modeling of speech has shown great potential for speech coding. However, the performance of such models drops when the input is not clean speech, e.g., in the presence of background noise, preventing its use in practical applications. In this paper we examine the reason and discuss methods to overcome this issue. Placing a denoising preprocessing stage when extracting features and target clean speech during training is shown to be the best performing strategy.
Explore related subjects
Keep this discovery
Tom Denton, Alejandro Luebs, Felicia S. C. Lim, Andrew Storus, Hengchin Yeh, W. Bastiaan Kleijn, Jan Skoglund. 2021-02-23. Handling Background Noise in Neural Speech Generation. https://arxiv.org/abs/2102.11906
Cite the original work for its findings. Save a collection to share your selection of sources.