arXiv · 2005.09252
Multi-Modal Summary Generation using Multi-Objective Optimization
Abstract
Significant development of communication technology over the past few years has motivated research in multi-modal summarization techniques. A majority of the previous works on multi-modal summarization focus on text and images. In this paper, we propose a novel extractive multi-objective optimization based model to produce a multi-modal summary containing text, images, and videos. Important objectives such as intra-modality salience, cross-modal redundancy and cross-modal similarity are optimized simultaneously in a multi-objective optimization framework to produce effective multi-modal output. The proposed model has been evaluated separately for different modalities, and has been found to perform better than state-of-the-art approaches.
Explore related subjects
Keep this discovery
Anubhav Jangra, Sriparna Saha, Adam Jatowt, Mohammad Hasanuzzaman. 2020-05-19. Multi-Modal Summary Generation using Multi-Objective Optimization. https://arxiv.org/abs/2005.09252
Cite the original work for its findings. Save a collection to share your selection of sources.