arXiv · 2609.30909
Machine Unlearning for Large Language Models: Foundations, Advances, and Agentic Extensions
Abstract
Machine unlearning aims to remove target influence while preserving other capabilities. This survey compares methods, benchmarks, and evidence across large language models and systems using retrieval, memory, tools, and interacting agents. A five-layer framework connects removal requests, system boundaries, target locations, interventions, and supported claims. A seven-stage lifecycle and six evidence dimensions guide comparison. The review shows that target construction, retained data, and recovery tests affect reported outcomes. Evidence from model evaluations remains insufficient to establish removal across external state and subsequent updates, motivating evaluation that tracks dependencies and tests whether target influence returns.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Xiaoyu Xu, Minxin Du, Li Bai, Junxu Liu, Yaxin Xiao, Kun Fang, Liu Yang, Huadi Zheng, Peizhao Hu, Qingqing Ye, Haibo Hu. 2026-09-25. Machine Unlearning for Large Language Models: Foundations, Advances, and Agentic Extensions. https://arxiv.org/abs/2609.30909
Cite the original work for its findings. Save a collection to share your selection of sources.