arXiv · 1811.03021
High-quality speech coding with SampleRNN
Abstract
We provide a speech coding scheme employing a generative model based on SampleRNN that, while operating at significantly lower bitrates, matches or surpasses the perceptual quality of state-of-the-art classic wide-band codecs. Moreover, it is demonstrated that the proposed scheme can provide a meaningful rate-distortion trade-off without retraining. We evaluate the proposed scheme in a series of listening tests and discuss limitations of the approach.
Explore related subjects
Keep this discovery
Janusz Klejsa, Per Hedelin, Cong Zhou, Roy Fejgin, Lars Villemoes. 2018-11-07. High-quality speech coding with SampleRNN. https://arxiv.org/abs/1811.03021
Cite the original work for its findings. Save a collection to share your selection of sources.