arXiv · 2503.13045
All You Need to Know About Training Image Retrieval Models
Abstract
Image retrieval is the task of finding images in a database that are most similar to a given query image. The performance of an image retrieval pipeline depends on many training-time factors, including the embedding model architecture, loss function, data sampler, mining function, learning rate(s), and batch size. In this work, we run tens of thousands of training runs to understand the effect each of these factors has on retrieval accuracy. We also discover best practices that hold across multiple datasets. The code is available at https://github.com/gmberton/image-retrieval
Explore related subjects
Keep this discovery
Gabriele Berton, Kevin Musgrave, Carlo Masone. 2025-03-17. All You Need to Know About Training Image Retrieval Models. https://arxiv.org/abs/2503.13045
Cite the original work for its findings. Save a collection to share your selection of sources.