PaperScope
LIVE · 2026-09-03 05:40 UTC

Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence

Eddie Conti, Claudio Daka, Álvaro Parafita, Antonio L. Alfeo, Axel Brando, Mario G. C. A. Cimino

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.00090 v1
Category
Submitted
2026-08-31

Abstract

Feature importance Methods (FIMs) are widely used in Explainable AI to interpret model predictions, yet attribution scores alone often provide limited insight into the underlying reasoning process. In this work, we introduce a novel perspective by embedding FIMs within a hypothesis-testing framework based on Weight of Evidence (WoE). We quantify how strongly the observed evidence supports any given hypothesis on feature importance. The reference hypothesis can stem from domain knowledge, ground truth, or be derived from the FIM itself. This formulation enables a principled evaluation of FIMs, capturing both their alignment with prior knowledge and their variability. We further provide theoretical results linking WoE to attribution variance. Empirical results shows the applicability and flexibility of our strategy analyzing LIME and SHAP explanations in settings with different reference hypotheses. Overall, our framework offers a complementary tool for assessing FIMs through a contrastive, evidence-based lens.

Comment: Accepted at XKDD and Beyond 2026 Workshop, ECML-PKDD

arXiv abs page · PDF · same-day batch