TY - RPRT TI - Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning AU - Antoine Yang AU - Arsha Nagrani AU - Paul Hongsuck Seo AU - Antoine Miech AU - Jordi Pont-Tuset AU - Ivan Laptev AU - Josef Sivic AU - Cordelia Schmid PY - 2023 UR - https://arxiv.org/abs/2302.14115 ID - 2302.14115 ER -