TY - RPRT TI - Aligning Information Capacity Between Vision and Language via Dense-to-Sparse Feature Distillation for Image-Text Matching AU - Yang Liu AU - Wentao Feng AU - Zhuoyao Liu AU - Shudong Huang AU - Jiancheng Lv PY - 2025 UR - https://arxiv.org/abs/2503.14953 ID - 2503.14953 ER -