TY - RPRT TI - DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding AU - Xiaoyi Bao AU - Chenwei Xie AU - Hao Tang AU - Tingyu Weng AU - Xiaofeng Wang AU - Yun Zheng AU - Xingang Wang PY - 2025 UR - https://arxiv.org/abs/2507.15569 ID - 2507.15569 ER -