TY - RPRT TI - VoxInstruct: Expressive Human Instruction-to-Speech Generation with Unified Multilingual Codec Language Modelling AU - Yixuan Zhou AU - Xiaoyu Qin AU - Zeyu Jin AU - Shuoyi Zhou AU - Shun Lei AU - Songtao Zhou AU - Zhiyong Wu AU - Jia Jia PY - 2024 DO - 10.1145/3664647.3681680 UR - https://arxiv.org/abs/2408.15676 ID - 2408.15676 ER -