arXiv Machine Learning

MinShap: A Shapley-Based Framework for Feature Redundancy

arXiv:2604. 15107v2 Announce Type: replace-cross Abstract: Shapley values provide a flexible framework for attributing feature contributions to model predictions, but they are not naturally suited for feature selection: a feature may receive a positive attribution even when it is redundant given the remaining variables.

arXiv Machine Learning
Jul 28

An Empirical Study of Feature Selection Granularity

arXiv:2607. 24145v1 Announce Type: new Abstract: Feature selection aims to identify the most informative and relevant features for a given dataset, either in terms of capturing the underlying data structure and distribution better, or with respect to the performance on a downstream task.

By Muhammad Rajabinasab, Arthur Zimek
arXiv Machine Learning
Jun 16

Priority-Aware Shapley Value

arXiv:2602. 09326v2 Announce Type: replace Abstract: Shapley values are widely used for model-agnostic data valuation and feature attribution, yet they implicitly assume contributors are interchangeable.

By Kiljae Lee, Ziqi Liu, Weijing Tang, Yuan Zhang
arXiv Machine Learning
Sep 18

Null importance: Disentangling relevance for interpretable machine learning

The paper introduces a unified framework called null importance to clarify different notions of feature relevance in interpretable machine learning. It defines null importance at the population level for various relevance concepts—marginal, conditional, predictive risk, functional invariance, and causal effects—and demonstrates how each answers distinct scientific questions. Through theoretical analysis, simulations, and case studies on fairness and genomic modeling, the authors show when these null notions coincide or diverge and how different importance methods target them.

By Garvesh Raskutti, Kris Sankaran, Jiaxin Ye
arXiv AI
Sep 2

Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence

The paper introduces a hypothesis‑testing framework that embeds feature importance methods (FIMs) within a Weight of Evidence (WoE) analysis. By quantifying how strongly observed evidence supports a given hypothesis—whether from domain knowledge, ground truth, or the FIM itself—the approach evaluates FIM alignment and variability. The authors provide theoretical links between WoE and attribution variance and demonstrate the method on LIME and SHAP explanations across varied reference hypotheses.

By Eddie Conti, Claudio Daka, \'Alvaro Parafita, Antonio L. Alfeo, Axel Brando, Mario G. C. A. Cimino
arXiv Machine Learning
Sep 18

Evaluating Explanation Methods by the Predictors They Induce

The paper proposes a new evaluation test for explanation methods: if an explanation accurately captures how a model uses its features, one should be able to reconstruct the model’s predictions from it. The authors convert explanations into predictors by summing feature effects and assess how well these predictors reproduce the model on unseen data, without any fitting. They apply this test to partial dependence plots, accumulated local effects, SHAP, and LIME across multiple datasets and model families, showing that the best method depends on feature dependence and that some existing quality metrics can favor flawed explanations.

By Jacob Selb{\ae}k, Hugo L. Hammer