Four Axiomatic Characterizations of the Integrated Gradients Attribution Method

06/23/2023
by   Daniel Lundstrom, et al.
0

Deep neural networks have produced significant progress among machine learning models in terms of accuracy and functionality, but their inner workings are still largely unknown. Attribution methods seek to shine a light on these "black box" models by indicating how much each input contributed to a model's outputs. The Integrated Gradients (IG) method is a state of the art baseline attribution method in the axiomatic vein, meaning it is designed to conform to particular principles of attributions. We present four axiomatic characterizations of IG, establishing IG as the unique method to satisfy different sets of axioms among a class of attribution methods.

READ FULL TEXT

Please sign up or login with your details

Forgot password? Click here to reset