Explainable AI Methods

Momentum

23 papers in the last four weeks, up 10% on the four weeks before. 0.3% of all new papers.

Jul 6Week of Sep 21

Latest papers 230

All topics
CardsList
  1. Concept-Based Abductive and Contrastive Explanations for Behaviors of Vision Models

    May 7, 2026Ronaldo Canizales, Divya Gopinath, Corina Păsăreanu +1ExplainabilityExplainable AI Methods

  2. eXplaining to Learn (eX2L): Regularization Using Contrastive Visual Explanation Pairs for Distribution Shifts

    May 7, 2026Paulo Mario P. Medina, Jose Marie Antonio Miñoza, Sebastian C. IbañezDistribution ShiftsContrastive Learning

  3. Explaining and Preventing Alignment Collapse in Iterative RLHF

    May 5, 2026Etienne Gauthier, Francis Bach, Michael I. JordanReinforcement Learning From Human FeedbackFrictive Policy Optimization

  4. Learning to Theorize the World from Observation

    May 5, 2026Doojin Baek, Gyubin Lee, Junyeob Baek +2Theory-Of-Mind ReasoningTheory

  5. Graph Reconstruction from Differentially Private GNN Explanations

    May 5, 2026Rishi Raj Sahoo, Jyotirmaya Shivottam, Subhankar MishraGraph Neural NetworksGradient-Based Attacks

  6. SAIL: Structure-Aware Interpretable Learning for Anatomy-Aligned Post-hoc Explanations in OCT

    May 4, 2026Tienyu Chang, Tianhao Li, Ruogu Fang +2Optical Coherence TomographyRetinal Imaging

  7. Less Interaction But More Explanation: A Communication Perspective on Agentic AI Interfaces

    May 2, 2026Eunchae Jang, S. Shyam SundarInteraction DataHuman Agency

  8. KG-First, LLM-Fallback: A Hybrid Microservice for Grounded Skill Search and Explanation

    May 2, 2026Ngoc Luyen Le, Marie-Hélène Abel, Bertrand LaforgeSkillsKnowledge Graphs

  9. Rethinking Explanations: Formalizing Contrast in Description Logics

    May 2, 2026Yasir Mahmood, Arnab Sharma, Axel-Cyrille Ngonga Ngomo +1Explainable AI MethodsAxiom

  10. LLMs Should Not Yet Be Credited with Decision Explanation

    May 1, 2026Wenshuo WangLarge Language Model DecisionsExplainable AI Methods

  11. Fairness of Classifiers in the Presence of Constraints between Features

    May 1, 2026Martin C. Cooper, Imane BousdiraAlgorithmic FairnessClassifier

  12. Are You the A-hole? A Fair, Multi-Perspective Ethical Reasoning Framework

    Apr 30, 2026Sheza Munir, Ahanaf Rodoshi, Sumin Lee +3EthicsDisagreement

  13. Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems

    Apr 28, 2026Ben Knight, Wm. Matthew Kennedy, Danielle Carvalho +2Large Language Models FailExplainable AI Methods

  14. Explanation of Dynamic Physical Field Predictions using WassersteinGrad: Application to Autoregressive Weather Forecasting

    Apr 24, 2026Younes Essafouri, Laure Raynaud, Luciano Drozda +1Artificial Intelligence Weather ModelsScalar Field

  15. Evaluating Post-hoc Explanations of the Transformer-based Genome Language Model DNABERT-2

    Apr 23, 2026Isabel Kurth, Paulo Yanez Sarmiento, Bernhard Y. RenardGenome-Wide Association StudyExplainable AI Methods

  16. Fine-Grained Perspectives: Modeling Explanations with Annotator-Specific Rationales

    Apr 23, 2026Olufunke O. Sarumi, Charles Welch, Daniel BraunHuman AnnotatorsExplainable AI Methods

  17. Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

    Apr 22, 2026Gricel Vázquez, Alexandros Evangelidis, Sepeedeh Shahbeigi +2Large Language Model PlanningAutomatic Prompt Optimization

  18. Concept Graph Convolutions: Message Passing in the Concept Space

    Apr 22, 2026Lucie Charlotte Magister, Pietro LioGraph Neural NetworksEdge-Aware

  19. PREF-XAI: Preference-Based Personalized Rule Explanations of Black-Box Machine Learning Models

    Apr 21, 2026Salvatore Greco, Jacek Karolczak, Roman Słowiński +1Explainable Artificial IntelligencePreference Learning

  20. From Top-1 to Top-K: A Reproducibility Study and Benchmarking of Counterfactual Explanations for Recommender Systems

    Apr 21, 2026Quang-Huy Nguyen, Thanh-Hai Nguyen, Khac-Manh Thai +6Counterfactual ExplanationExplainable AI Methods

  21. TACENR: Task-Agnostic Contrastive Explanations for Node Representations

    Apr 21, 2026Vasiliki Papanikou, Evaggelia PitouraGraph Representation LearningExplainable AI Methods

  22. From Scoring to Explanations: Evaluating SHAP and LLM Rationales for Rubric-based Teaching Quality Assessment

    Apr 18, 2026Ivo Bueno, Babette Bühler, Philipp Stark +5Rubric-Based ScoringFree-Form Textual Rationales

  23. The Query Channel: Information-Theoretic Limits of Masking-Based Explanations

    Apr 17, 2026Erciyes Karakaya, Ozgur ErcetinInformation-Theoretic LimitsExplainable AI Methods

  24. Explain the Flag: Contextualizing Hate Speech Beyond Censorship

    Apr 16, 2026Jason Liartis, Eirini Kaldeli, Lambrini Gyftokosta +2Hate Speech DetectionHate Speech

  25. Rethinking Patient Education as Multi-turn Multi-modal Interaction

    Apr 16, 2026Zonghai Yao, Zhipeng Tang, Chengtao Lin +5Multimodal Clinical DataMedical Visual Question Answering

  26. Robust Explanations for User Trust in Enterprise NLP Systems

    Apr 13, 2026Guilin Zhang, Kai Zhao, Jeffrey Friedman +3Large Language Model ReliabilityExplainability

  27. Exploring Plan Space through Conversation: An Agentic Framework for LLM-Mediated Explanations in Planning

    Mar 2, 2026Guilhem Fouilhé, Rebecca Eifler, Antonin Poché +2Large Language Model PlanningAgentic Framework

  28. Quantifying Retriever-Generator Alignment in RAG with Local Explanations

    Jan 29, 2026Korbinian Randl, Guido Rocchietti, Aron Henriksson +3Agentic Retrieval-Augmented Generation SystemsExplainability

  29. TxSum: User-Centered Ethereum Transaction Understanding with Micro-Level Semantic Grounding

    Dec 7, 2025Zifan Peng, Jingyi Zheng, Yule Liu +8BitcoinTransaction Evidence

  30. Critical or Compliant? The Double-Edged Sword of Reasoning in Chain-of-Thought Explanations

    Nov 15, 2025Eunkyu Park, Wesley Hanwen Deng, Vasudha Varadarajan +4Chain-of-Thought ReasoningReasoning Errors

  31. Multi-Step Knowledge Interaction Analysis via Rank-2 Subspace Disentanglement

    Nov 3, 2025Sekh Mainul Islam, Pepa Atanasova, Isabelle AugensteinLLM Reasoning StrategiesContextual Grounding

  32. Activation-Deactivation: A General Framework for Robust Post-hoc Explainable AI

    Oct 1, 2025Akchunya Chanchal, David A. Kelly, Hana ChocklerExplainabilityExplainable Artificial Intelligence

  33. Cross-Attention is Half Explanation in Speech-to-Text Models

    Sep 22, 2025Sara Papi, Dennis Fucci, Marco Gaido +2Cross-AttentionNatural Language Processing

  34. Ultra Strong Machine Learning: LLM-Generated Explanations Do Not Yet Suffice for Teaching Humans Active Learning Strategy

    Aug 31, 2025Lun Ai, Johannes Langer, Ute Schmid +1Active LearningInteractive Learning

  35. Explain Before You Answer: A Survey on Compositional Visual Reasoning

    Aug 24, 2025Fucai Ke, Joy Hsu, Zhixi Cai +10Visual ReasoningMultimodal Reasoning

  36. Rule2Text: A Framework for Generating and Evaluating Natural Language Explanations of Knowledge Graph Rules

    Aug 14, 2025Nasim Shirvani-Mahdavi, Chengkai LiKnowledge GraphsExplainable AI Methods

  37. A Communication-First Account of Explanation

    May 6, 2025Jacqueline Harding, Tobias Gerstenberg, Thomas IcardExplainable AI MethodsCausal

  38. Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations

    Apr 7, 2025Pedro Ferreira, Wilker Aziz, Ivan TitovExplainable AI MethodsChain-of-Thought Reasoning

  39. Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

    Mar 17, 2025Noah Y. Siegel, Nicolas Heess, Maria Perez-Ortiz +1FaithfulnessExplainable AI Methods

  40. Robust Counterfactual Explanations under Model Multiplicity Using Multi-Objective Optimization

    Jan 10, 2025Keita KinjoCounterfactual ExplanationExplainability

  41. Causal Explanations for Image Classifiers

    Nov 13, 2024Hana Chockler, David A. Kelly, Daniel Kroening +1ExplainabilityExplainable AI Methods

  42. Towards Complete Causal Explanation with Expert Knowledge

    Jul 10, 2024Aparajithan Venkateswaran, Emilija PerkovićAcyclic GraphsCausal

  43. Dont Just Teach, Explain! A Gamified 20Q Recommender for Cybersecurity Education

    Date pendingMary Nusrat, Sarfuddin Bhuiyan, Gahangir HossainCybersecurityInteractive Learning

  44. SAEExplainer: Interpreting SAE Features with Activation-Guided Preference Optimization

    Date pendingJingyi He, Haiyan Zhao, Ruxue Shi +4Sparse Autoencoder FeaturesExplainable AI Methods

  45. Limitations of Automated Simulatability: LLM Simulators Can Bypass Explanations

    Date pendingAntonin Poché, Fanny Jourdan, Nils Feldhus +6User SimulationExplainable AI Methods

  46. PACE: A Neuro-Symbolic Framework for Plausible and Actionable Counterfactual Explanations

    Date pendingPavel Iakovets, Liyanapathiranage Sudeepika Wajirakumari Samarathunga, Martin Thomas Horsch +1Counterfactual ExplanationExplainable Artificial Intelligence