UbiQVision: Quantifying Uncertainty in XAI for Image Recognition
cs.CV, cs.AI
Submitted: 2025-12-23
Updated: 2026-04-23
Comments: Under Review. Updated manuscript. Feedback from reviewers incorporated
DOI: 10.1016/j.mlwa.2026.101000
License: http://creativecommons.org/licenses/by-nc-sa/4.0/
The gist: Recent advances in deep learning have led to its widespread adoption across diverse domains, including medical imaging.
Terminology
Abstract
Recent advances in deep learning have led to its widespread adoption across diverse domains, including medical imaging. This progress is driven by increasingly sophisticated model architectures, such as ResNets, Vision Transformers, and Hybrid Convolutional Neural Networks, that offer enhanced performance at the cost of greater complexity. This complexity often compromises model explainability and interpretability. SHAP has emerged as a prominent method for providing interpretable visualizations that aid domain experts in understanding model predictions. However, SHAP explanations can be unstable and unreliable in the presence of epistemic and aleatoric uncertainty. In this study, we address this challenge by using Dirichlet posterior sampling and Dempster-Shafer theory to quantify the uncertainty that arises from these unstable explanations in medical imaging applications. The framework uses a belief, plausible, and fusion map approach alongside statistical quantitative analysis to produce quantification of uncertainty in SHAP. Furthermore, we evaluated our framework on three medical imaging datasets with varying class distributions, image qualities, and modality types which introduces noise due to varying image resolutions and modality-specific aspect covering the examples from pathology, ophthalmology, and radiology, introducing significant epistemic uncertainty.
Sources
- Evaluating the Explainability of Vision Transformers in Medical Imaging
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- Deep Learning Applications in Medical Image Analysis: Advancements, Challenges, and Future Directions
- Axiom-based Grad-CAM: Towards Accurate Visualization and Explanation of CNNs
- Human Agency, Causality, and the Human Computer Interface in High-Stakes Artificial Intelligence
- On Uncertainty, Tempering, and Data Augmentation in Bayesian Classification
- Here Comes the Explanation: A Shapley Perspective on Multi-contrast Medical Image Segmentation
- Probabilistic Lipschitzness and the Stable Rank for Comparing Explanation Models
- Tempering the Bayes Filter towards Improved Model-Based Estimation
- Explaining Predictive Uncertainty with Information Theoretic Shapley Values
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models