arXiv · 2601.10936
Can Instructed Retrieval Models Really Support Exploration?
Abstract
Exploratory searches are characterized by under-specified goals and evolving query intents. In such scenarios, retrieval models that can capture user-specified nuances in query intent and adapt results accordingly are desirable -- instruction-following retrieval models promise such a capability. In this work, we evaluate instructed retrievers for the prevalent yet under-explored application of aspect-conditional seed-guided exploration using an expert-annotated test collection. We evaluate both recent LLMs fine-tuned for instructed retrieval and general-purpose LLMs prompted for ranking with the highly performant Pairwise Ranking Prompting. We find that the best instructed retrievers improve on ranking relevance compared to instruction-agnostic approaches. However, we also find that instruction following performance, crucial to the user experience of interacting with models, does not mirror ranking relevance improvements and displays insensitivity or counter-intuitive behavior to instructions. Our results indicate that while users may benefit from using current instructed retrievers over instruction-agnostic models, they may not benefit from using them for long-running exploratory sessions requiring greater sensitivity to instructions.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Piyush Maheshwari, Sheshera Mysore, Hamed Zamani. 2026-01-16. Can Instructed Retrieval Models Really Support Exploration?. https://doi.org/10.1145/3786304.3787888
Cite the original work for its findings. Save a collection to share your selection of sources.