TY - RPRT TI - UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts AU - Zhi-Qi Cheng AU - Xiang Li AU - Jun-Yan He AU - Junyao Chen AU - Xiaomao Fan AU - Xiaojiang Peng AU - Alexander G. Hauptmann PY - 2025 UR - https://arxiv.org/abs/2404.18398 ID - 2404.18398 ER -