TY - RPRT TI - Unveiling the Potential of Vision-Language-Action Models with Open-Ended Multimodal Instructions AU - Wei Zhao AU - Gongsheng Li AU - Zhefei Gong AU - Pengxiang Ding AU - Han Zhao AU - Donglin Wang PY - 2025 UR - https://arxiv.org/abs/2505.11214 ID - 2505.11214 ER -