TY - RPRT TI - TokenFocus-VQA: Enhancing Text-to-Image Alignment with Position-Aware Focus and Multi-Perspective Aggregations on LVLMs AU - Zijian Zhang AU - Xuhui Zheng AU - Xuecheng Wu AU - Chong Peng AU - Xuezhi Cao PY - 2025 UR - https://arxiv.org/abs/2504.07556 ID - 2504.07556 ER -