TY - RPRT TI - ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts AU - Mu Cai AU - Haotian Liu AU - Dennis Park AU - Siva Karthik Mustikovela AU - Gregory P. Meyer AU - Yuning Chai AU - Yong Jae Lee PY - 2024 UR - https://arxiv.org/abs/2312.00784 ID - 2312.00784 ER -