arXiv · 1907.03077
Generative Counterfactual Introspection for Explainable Deep Learning
Abstract
In this work, we propose an introspection technique for deep neural networks that relies on a generative model to instigate salient editing of the input image for model interpretation. Such modification provides the fundamental interventional operation that allows us to obtain answers to counterfactual inquiries, i.e., what meaningful change can be made to the input image in order to alter the prediction. We demonstrate how to reveal interesting properties of the given classifiers by utilizing the proposed introspection approach on both the MNIST and the CelebA dataset.
Explore related subjects
Keep this discovery
Shusen Liu, Bhavya Kailkhura, Donald Loveland, Yong Han. 2019-07-06. Generative Counterfactual Introspection for Explainable Deep Learning. https://arxiv.org/abs/1907.03077
Cite the original work for its findings. Save a collection to share your selection of sources.