arXiv · 2505.09660
On Measuring Intrinsic Causal Attributions in Deep Neural Networks
Abstract
Quantifying the causal influence of input features within neural networks has become a topic of increasing interest. Existing approaches typically assess direct, indirect, and total causal effects. This work treats NNs as structural causal models (SCMs) and extends our focus to include intrinsic causal contributions (ICC). We propose an identifiable generative post-hoc framework for quantifying ICC. We also draw a relationship between ICC and Sobol' indices. Our experiments on synthetic and real-world datasets demonstrate that ICC generates more intuitive and reliable explanations compared to existing global explanation techniques.
Explore related subjects
Keep this discovery
Saptarshi Saha, Dhruv Vansraj Rathore, Soumadeep Saha, Utpal Garain, David Doermann. 2025-05-14. On Measuring Intrinsic Causal Attributions in Deep Neural Networks. https://arxiv.org/abs/2505.09660
Cite the original work for its findings. Save a collection to share your selection of sources.