On the Instability of Saliency Maps under Sparse Perturbations
Understanding the reliability of model explanations remains a critical challenge in deep learning. Prior work has shown that saliency maps can be manipulated by optimizing the input using gradient-based methods, where gradients of the loss with respect to the input are computed to generate dense perturbations that alte...