arXiv · 2412.00666
Explaining Object Detectors via Collective Contribution of Pixels
Abstract
Visual explanations for object detectors are crucial for enhancing their reliability. Object detectors identify and localize instances by assessing multiple visual features collectively. When generating explanations, overlooking these collective influences in detections may lead to missing compositional cues or capturing spurious correlations. However, existing methods typically focus solely on individual pixel contributions, neglecting the collective contribution of multiple pixels. To address this limitation, we propose a game-theoretic method based on Shapley values and interactions to explicitly capture both individual and collective pixel contributions. Our method provides explanations for both bounding box localization and class determination, highlighting regions crucial for detection. Extensive experiments demonstrate that the proposed method identifies important regions more accurately than state-of-the-art methods. The code is available at https://github.com/tttt-0814/VX-CODE
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Toshinori Yamauchi, Hiroshi Kera, Kazuhiko Kawamoto. 2024-12-01. Explaining Object Detectors via Collective Contribution of Pixels. https://arxiv.org/abs/2412.00666
Cite the original work for its findings. Save a collection to share your selection of sources.