TY - RPRT TI - COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning AU - Simon Ging AU - Mohammadreza Zolfaghari AU - Hamed Pirsiavash AU - Thomas Brox PY - 2020 UR - https://arxiv.org/abs/2011.00597 ID - 2011.00597 ER -