2025
Explain Yourself, Briefly! Self-Explaining Neural Networks with Concise Sufficient Reasons
ICLR 2025poster
*Minimal sufficient reasons* represent a prevalent form of explanation - the smallest subset of input features which, when held constant at their corresponding values, ensure that the prediction remains unchanged. Previous *post-hoc* methods attempt to obtain such explanations but face two main limi…