Four Axiomatic Characterizations of the Integrated Gradients Attribution Method (2306.13753v1)

Published 23 Jun 2023 in cs.LG

Abstract: Deep neural networks have produced significant progress among machine learning models in terms of accuracy and functionality, but their inner workings are still largely unknown. Attribution methods seek to shine a light on these "black box" models by indicating how much each input contributed to a model's outputs. The Integrated Gradients (IG) method is a state of the art baseline attribution method in the axiomatic vein, meaning it is designed to conform to particular principles of attributions. We present four axiomatic characterizations of IG, establishing IG as the unique method to satisfy different sets of axioms among a class of attribution methods.

References (35)

Citations (2)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Four Axiomatic Characterizations of the Integrated Gradients Attribution Method (2306.13753v1)

Summary

Related Papers