arXiv · 1710.11473
Multi-Resolution Fully Convolutional Neural Networks for Monaural Audio Source Separation
Abstract
In deep neural networks with convolutional layers, each layer typically has fixed-size/single-resolution receptive field (RF). Convolutional layers with a large RF capture global information from the input features, while layers with small RF size capture local details with high resolution from the input features. In this work, we introduce novel deep multi-resolution fully convolutional neural networks (MR-FCNN), where each layer has different RF sizes to extract multi-resolution features that capture the global and local details information from its input features. The proposed MR-FCNN is applied to separate a target audio source from a mixture of many audio sources. Experimental results show that using MR-FCNN improves the performance compared to feedforward deep neural networks (DNNs) and single resolution deep fully convolutional neural networks (FCNNs) on the audio source separation problem.
Explore related subjects
Keep this discovery
Emad M. Grais, Hagen Wierstorf, Dominic Ward, Mark D. Plumbley. 2017-10-28. Multi-Resolution Fully Convolutional Neural Networks for Monaural Audio Source Separation. https://arxiv.org/abs/1710.11473
Cite the original work for its findings. Save a collection to share your selection of sources.