arXiv · 2209.02785
Read it to me: An emotionally aware Speech Narration Application
Abstract
In this work we try to perform emotional style transfer on audios. In particular, MelGAN-VC architecture is explored for various emotion-pair transfers. The generated audio is then classified using an LSTM-based emotion classifier for audio. We find that "sad" audio is generated well as compared to "happy" or "anger" as people have similar expressions of sadness.
Explore related subjects
Keep this discovery
Rishibha Bansal. 2022-09-06. Read it to me: An emotionally aware Speech Narration Application. https://arxiv.org/abs/2209.02785
Cite the original work for its findings. Save a collection to share your selection of sources.