arXiv · 2108.11637
Self-Attention for Audio Super-Resolution
Abstract
Convolutions operate only locally, thus failing to model global interactions. Self-attention is, however, able to learn representations that capture long-range dependencies in sequences. We propose a network architecture for audio super-resolution that combines convolution and self-attention. Attention-based Feature-Wise Linear Modulation (AFiLM) uses self-attention mechanism instead of recurrent neural networks to modulate the activations of the convolutional model. Extensive experiments show that our model outperforms existing approaches on standard benchmarks. Moreover, it allows for more parallelization resulting in significantly faster training.
Explore related subjects
Keep this discovery
Nathanaël Carraz Rakotonirina. 2021-08-26. Self-Attention for Audio Super-Resolution. https://arxiv.org/abs/2108.11637
Cite the original work for its findings. Save a collection to share your selection of sources.