cs.LGAug 31, 2026

Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence

Authors: Eddie ContiClaudio DakaÁlvaro ParafitaAntonio L. AlfeoAxel BrandoMario G. C. A. Cimino

Organizations: Barcelona Supercomputing Center, Barcelona, Spain · University of Florence, Florence, Italy · SMARTEST Research Center, eCampus University, Novedrate, Italy · Dept. of Theoretical and Applied Sciences, eCampus University, Novedrate, Italy · Dept Information Engineering, University of Pisa, Pisa, Italy

Abstract

Feature importance Methods (FIMs) are widely used in Explainable AI to interpret model predictions, yet attribution scores alone often provide limited insight into the underlying reasoning process. In this work, we introduce a novel perspective by embedding FIMs within a hypothesis-testing framework based on Weight of Evidence (WoE). We quantify how strongly the observed evidence supports any given hypothesis on feature importance. The reference hypothesis can stem from domain knowledge, ground truth, or be derived from the FIM itself. This formulation enables a principled evaluation of FIMs, capturing both their alignment with prior knowledge and their variability. We further provide theoretical results linking WoE to attribution variance. Empirical results shows the applicability and flexibility of our strategy analyzing LIME and SHAP explanations in settings with different reference hypotheses. Overall, our framework offers a complementary tool for assessing FIMs through a contrastive, evidence-based lens.

Explore similar work

CardsList
  1. Measuring Explainer Stability via Attribution Separability

    Aug 3, 2026Eddie Conti, Álvaro Parafita, Axel BrandoExplainabilityBlack Box