TY - RPRT TI - Voice Activity Projection Model with Multimodal Encoders AU - Takeshi Saga AU - Catherine Pelachaud PY - 2025 UR - https://arxiv.org/abs/2506.03980 ID - 2506.03980 ER -