arXiv · 2501.03989
(De)-Indexing and the Right to be Forgotten
Abstract
In the digital age, the challenge of forgetfulness has emerged as a significant concern, particularly regarding the management of personal data and its accessibility online. The right to be forgotten (RTBF) allows individuals to request the removal of outdated or harmful information from public access, yet implementing this right poses substantial technical difficulties for search engines. This paper aims to introduce non-experts to the foundational concepts of information retrieval (IR) and de-indexing, which are critical for understanding how search engines can effectively "forget" certain content. We will explore various IR models, including boolean, probabilistic, vector space, and embedding-based approaches, as well as the role of Large Language Models (LLMs) in enhancing data processing capabilities. By providing this overview, we seek to highlight the complexities involved in balancing individual privacy rights with the operational challenges faced by search engines in managing information visibility.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Salvatore Vilella, Giancarlo Ruffo. 2025-01-07. (De)-Indexing and the Right to be Forgotten. https://arxiv.org/abs/2501.03989
Cite the original work for its findings. Save a collection to share your selection of sources.