TY - RPRT TI - Scalable and Accurate Self-supervised Multimodal Representation Learning without Aligned Video and Text Data AU - Vladislav Lialin AU - Stephen Rawls AU - David Chan AU - Shalini Ghosh AU - Anna Rumshisky AU - Wael Hamza PY - 2023 DO - 10.1109/wacvw58289.2023.00043 UR - https://arxiv.org/abs/2304.02080 ID - 2304.02080 ER -