TY - RPRT TI - SynthVLM: Towards High-Quality and Efficient Synthesis of Image-Caption Datasets for Vision-Language Models AU - Zheng Liu AU - Hao Liang AU - Bozhou Li AU - Wentao Xiong AU - Chong Chen AU - Conghui He AU - Wentao Zhang AU - Bin Cui PY - 2025 UR - https://arxiv.org/abs/2407.20756 ID - 2407.20756 ER -