arXiv · 2204.02524
Simple and Effective Unsupervised Speech Synthesis
Abstract
We introduce the first unsupervised speech synthesis system based on a simple, yet effective recipe. The framework leverages recent work in unsupervised speech recognition as well as existing neural-based speech synthesis. Using only unlabeled speech audio and unlabeled text as well as a lexicon, our method enables speech synthesis without the need for a human-labeled corpus. Experiments demonstrate the unsupervised system can synthesize speech similar to a supervised counterpart in terms of naturalness and intelligibility measured by human evaluation.
Explore related subjects
Keep this discovery
Alexander H. Liu, Cheng-I Jeff Lai, Wei-Ning Hsu, Michael Auli, Alexei Baevski, James Glass. 2022-04-06. Simple and Effective Unsupervised Speech Synthesis. https://arxiv.org/abs/2204.02524
Cite the original work for its findings. Save a collection to share your selection of sources.