IViT: A Novel Interpretable Visual Transformer for Skin Disease Detection
eess.IV, cs.CV
Submitted: 2026-06-22
Updated: 2026-09-22
Terminology
Sources
- Skin Lesion Analysis Toward Melanoma Detection 2018: A Challenge Hosted by the International Skin Imaging Collaboration (ISIC)
- Deep Residual Learning for Image Recognition
- Visualizing and Understanding Convolutional Networks
- Attention Is All You Need
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- Visual Bias and Interpretability in Deep Learning for Dermatological Image Analysis
- MT-TransUNet: Mediating Multi-Task Tokens in Transformers for Skin Lesion Segmentation and Classification
- Transformer Interpretability Beyond Attention Visualization
- "Why Should I Trust You?": Explaining the Predictions of Any Classifier
- Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
- Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers
Related papers
- Revisiting Integration of Image and Metadata for DICOM Series Classification: Cross-Attention and Dictionary Learning
- VesselSDF: Distance Field Priors for Vascular Network Reconstruction
- cSVR: Convolutional Slice-to-Volume Reconstruction
- NAIMA: Semantics Aware RGB Guided Depth Super-Resolution
- AneumoBench: A Source-Linked Benchmark for Synthetic-Geometry Transfer in Aneurysm CFD
- RETO: A Rotary-Enhanced Transformer Operator for High-Fidelity Prediction of Automotive Aerodynamics