arXiv · 2409.10907
Attention-Seeker: Dynamic Self-Attention Scoring for Unsupervised Keyphrase Extraction
Abstract
This paper proposes Attention-Seeker, an unsupervised keyphrase extraction method that leverages self-attention maps from a Large Language Model to estimate the importance of candidate phrases. Our approach identifies specific components - such as layers, heads, and attention vectors - where the model pays significant attention to the key topics of the text. The attention weights provided by these components are then used to score the candidate phrases. Unlike previous models that require manual tuning of parameters (e.g., selection of heads, prompts, hyperparameters), Attention-Seeker dynamically adapts to the input text without any manual adjustments, enhancing its practical applicability. We evaluate Attention-Seeker on four publicly available datasets: Inspec, SemEval2010, SemEval2017, and Krapivin. Our results demonstrate that, even without parameter tuning, Attention-Seeker outperforms most baseline models, achieving state-of-the-art performance on three out of four datasets, particularly excelling in extracting keyphrases from long documents.
Explore related subjects
Keep this discovery
Erwin D. López Z., Cheng Tang, Atsushi Shimada. 2024-09-17. Attention-Seeker: Dynamic Self-Attention Scoring for Unsupervised Keyphrase Extraction. https://arxiv.org/abs/2409.10907
Cite the original work for its findings. Save a collection to share your selection of sources.