arXiv · 2407.03610
VDMA: Video Question Answering with Dynamically Generated Multi-Agents
Abstract
This technical report provides a detailed description of our approach to the EgoSchema Challenge 2024. The EgoSchema Challenge aims to identify the most appropriate responses to questions regarding a given video clip. In this paper, we propose Video Question Answering with Dynamically Generated Multi-Agents (VDMA). This method is a complementary approach to existing response generation systems by employing a multi-agent system with dynamically generated expert agents. This method aims to provide the most accurate and contextually appropriate responses. This report details the stages of our approach, the tools employed, and the results of our experiments.
Explore related subjects
Keep this discovery
Noriyuki Kugo, Tatsuya Ishibashi, Kosuke Ono, Yuji Sato. 2024-07-04. VDMA: Video Question Answering with Dynamically Generated Multi-Agents. https://arxiv.org/abs/2407.03610
Cite the original work for its findings. Save a collection to share your selection of sources.