TY - RPRT TI - Vision Model Pre-training on Interleaved Image-Text Data via Latent Compression Learning AU - Chenyu Yang AU - Xizhou Zhu AU - Jinguo Zhu AU - Weijie Su AU - Junjie Wang AU - Xuan Dong AU - Wenhai Wang AU - Lewei Lu AU - Bin Li AU - Jie Zhou AU - Yu Qiao AU - Jifeng Dai PY - 2024 UR - https://arxiv.org/abs/2406.07543 ID - 2406.07543 ER -