arXiv · 1102.3865
Probability Based Clustering for Document and User Properties
Abstract
Information Retrieval systems can be improved by exploiting context information such as user and document features. This article presents a model based on overlapping probabilistic or fuzzy clusters for such features. The model is applied within a fusion method which linearly combines several retrieval systems. The fusion is based on weights for the different retrieval systems which are learned by exploiting relevance feedback information. This calculation can be improved by maintaining a model for each document and user cluster. That way, the optimal retrieval system for each document or user type can be identified and applied. The extension presented in this article allows overlapping, probabilistic clusters of features to further refine the process.
Explore related subjects
Keep this discovery
Thomas Mandl, Christa Womser-Hacker. 2011-02-18. Probability Based Clustering for Document and User Properties. https://arxiv.org/abs/1102.3865
Cite the original work for its findings. Save a collection to share your selection of sources.