Investigating Saturation Effects in Integrated Gradients (2010.12697v1)

Published 23 Oct 2020 in cs.CV and cs.LG

Abstract: Integrated Gradients has become a popular method for post-hoc model interpretability. De-spite its popularity, the composition and relative impact of different regions of the integral path are not well understood. We explore these effects and find that gradients in saturated regions of this path, where model output changes minimally, contribute disproportionately to the computed attribution. We propose a variant of IntegratedGradients which primarily captures gradients in unsaturated regions and evaluate this method on ImageNet classification networks. We find that this attribution technique shows higher model faithfulness and lower sensitivity to noise com-pared with standard Integrated Gradients. A note-book illustrating our computations and results is available at https://github.com/vivekmig/captum-1/tree/ExpandedIG.

Authors (5)

Vivek Miglani (7 papers)
Narine Kokhlikyan (15 papers)
Bilal Alsallakh (11 papers)
Miguel Martin (53 papers)
Orion Reblitz-Richardson (5 papers)

Citations (22)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Investigating Saturation Effects in Integrated Gradients (2010.12697v1)

Summary

Related Papers

GitHub