arXiv · 2012.15483
Why do classifier accuracies show linear trends under distribution shift?
Abstract
Recent studies of generalization in deep learning have observed a puzzling trend: accuracies of models on one data distribution are approximately linear functions of the accuracies on another distribution. We explain this trend under an intuitive assumption on model similarity, which was verified empirically in prior work. More precisely, we assume the probability that two models agree in their predictions is higher than what we can infer from their accuracy levels alone. Then, we show that a linear trend must occur when evaluating models on two distributions unless the size of the distribution shift is large. This work emphasizes the value of understanding model similarity, which can have an impact on the generalization and robustness of classification models.
Explore related subjects
Keep this discovery
Horia Mania, Suvrit Sra. 2020-12-31. Why do classifier accuracies show linear trends under distribution shift?. https://arxiv.org/abs/2012.15483
Cite the original work for its findings. Save a collection to share your selection of sources.