TY - RPRT TI - Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs AU - Ling Yang AU - Zhaochen Yu AU - Chenlin Meng AU - Minkai Xu AU - Stefano Ermon AU - Bin Cui PY - 2024 UR - https://arxiv.org/abs/2401.11708 ID - 2401.11708 ER -