arXiv:2609.36771v1 Announce Type: cross
Abstract: Root cause analysis (RCA) is a critical problem in many real-world scenarios. RCA enables the identification of faulty or failing mechanisms in a sys...
By Md Musfiqur Rahman, Kenneth Lee, Ziwei Jiang, Padmaja Jonnalagedda, Ruocheng Guo, Murat Kocaoglu
arXiv:2610.00968v1 Announce Type: cross
Abstract: Causal representation learning aims to discover robust features by exploiting the causal structure underlying data generation. Existing methods requi...
By Arman Behnam, Binghui Wang
The paper introduces Probabilistic Causal Impact (PCI), a framework that blends actual causality (AC) with Pearl’s probability of necessity and sufficiency to provide tractable, causally grounded explanations. PCI reframes explainability as an estimation problem on a probabilistic causal model, enabling efficient approximation via Monte Carlo sampling. The authors evaluate PCI on synthetic and real-world data, demonstrating consistency with AC, scalability, and applicability to complex continuous systems and large-scale causal machine learning models.
By Rafal Urbaniak, Sam Witty, Daniel Waxman, Andy Zane, Poorva Garg, Emily Bunnapradist, Sankaran Vaidyanathan, Jack Feser, Drew Lehe, Eli Bingham
arXiv:2608.23835v1 Announce Type: new
Abstract: Real-world decision-making in public health and social science can greatly benefit from predictive models, yet translating predictions into effective i...
By Mulin Tian, Ajitesh Srivastava
arXiv:2602. 08629v2 Announce Type: replace Abstract: Causal discovery is essential for advancing data-driven fields such as scientific AI and data analysis, yet existing approaches face significant time- and space-efficiency bottlenecks when scaling to large graphs.
By Bo Peng, Sirui Chen, Jiaguo Tian, Yu Qiao, Chaochao Lu
arXiv:2604. 27007v2 Announce Type: replace Abstract: We provide a causal analysis of Binary Spiking Neural Networks (BSNNs) to explain their behavior.
By Aditya Kar (CNRS, IRIT), Emiliano Lorini (CNRS, IRIT), Timoth\'ee Masquelier (CNRS, CERCO UMR5549)
arXiv:2411. 08875v4 Announce Type: replace Abstract: Existing algorithms for explaining the output of image classifiers use different definitions of explanations and a variety of techniques to find them.
By Hana Chockler, David A. Kelly, Daniel Kroening, Youcheng Sun
arXiv:2602. 06337v2 Announce Type: replace-cross Abstract: Causal inference is essential for decision-making but remains challenging for non-experts.
By Junqi Chen, Sirui Chen, Chaochao Lu
arXiv:2404.06349v3 Announce Type: replace
Abstract: The ability to understand causality significantly impacts the competence of large language models (LLMs) in output explanation and counterfactual r...
By Yu Zhou, Xingyu Wu, Jibin Wu, Liang Feng, Kay Chen Tan
arXiv:2510. 27544v3 Announce Type: replace Abstract: Current training paradigms, optimized for long-horizon reasoning trace execution, have made Large Language Models (LLMs) excel at pattern matching and forward simulation of reasoning, but underperform at counterfactual causal understanding and reasoning.
By Nikolaus Holzer, William Fishell, Baishakhi Ray, Mark Santolucito
arXiv:2607. 08641v1 Announce Type: new Abstract: Over the last few years, there has been an increased interest in making machine learning models more interpretable.
By Yann Claes, Pierre Geurts, V\^an Anh Huynh-Thu
Matryoshka Attribution (MAttr) is a mask‑learning method that identifies nested subsets of a language model’s internal components by minimizing downstream loss. It uses a differentiable sigmoid top‑k operator and randomizes sparsity during training to produce an attribution ordering of components. MAttr tops the Mechanistic Interpretability Benchmark leaderboard and can be applied via reinforcement learning to pinpoint weight changes that control behaviors such as refusal in Llama 3.1 8B Instruct, where restoring just 1% of weights removes refusals while preserving capabilities.
By Aryaman Arora, Kirill Acharya, Nathan Hu, Yanzhe Zhang, Noah Goodman, Dan Jurafsky, Christopher Potts