TY - RPRT TI - DiVISe: Direct Visual-Input Speech Synthesis Preserving Speaker Characteristics And Intelligibility AU - Yifan Liu AU - Yu Fang AU - Zhouhan Lin PY - 2025 UR - https://arxiv.org/abs/2503.05223 ID - 2503.05223 ER -