arXiv · 2609.40121
On the (In)effectiveness of AMR Augmentation for Large Language Models
Abstract
While Abstract Meaning Representation (AMR) has historically improved performance on a range of NLP tasks, the benefit---or lack thereof---of AMR augmentation for modern LLMs is thus far unclear. In this paper, we attempt to reproduce recent work that reported substantial downstream gains from AMR augmentation, finding that these are likely due to specific choices in the experimental settings used: using a consistent and unified protocol for hyperparameter selection, we observe that text-only baselines consistently match or exceed the performance of AMR-augmented models. To investigate this null result, we introduce a perplexity-based probe measuring the degree to which AMR provides an LLM with supplemental relational knowledge not already available to the model. We find that AMR augmentation does not help LLMs improve their understanding of relational content in the sentence, indicating that augmenting these models with AMR offers no clear benefit on downstream tasks.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hoa Quynh Nhung Nguyen, Jacopo Staiano, Michael Sullivan. 2026-09-30. On the (In)effectiveness of AMR Augmentation for Large Language Models. https://arxiv.org/abs/2609.40121
Cite the original work for its findings. Save a collection to share your selection of sources.