TY - RPRT TI - Do Egocentric Video-Language Models Capture Both Hand- and Object-Centric Cues? AU - Masatoshi Tateno AU - Alexandros Stergiou AU - Risa Shinoda AU - Yoichi Sato AU - Dima Damen PY - 2026 UR - https://arxiv.org/abs/2607.08514 ID - 2607.08514 ER -