TY - RPRT TI - Scaling Vision-Language Models with Sparse Mixture of Experts AU - Sheng Shen AU - Zhewei Yao AU - Chunyuan Li AU - Trevor Darrell AU - Kurt Keutzer AU - Yuxiong He PY - 2023 UR - https://arxiv.org/abs/2303.07226 ID - 2303.07226 ER -