arXiv Computer Vision

IViT: A Novel Interpretable Visual Transformer for Skin Disease Detection

Hugging Face Trending Papers
Aug 11

Uncertainty-Aware and Explainable Ensemble Deep Learning Framework for Multi-Class Skin Lesion Classification

Skin cancer diagnosis from dermoscopic images remains challenging due to high intra-class variability, inter-class similarity, class imbalance, and the limited interpretability of deep learning models. This paper proposes an uncertainty-aware and explainable deep learning framework for multi-class skin lesion classification.

arXiv AI
Sep 3

Disease Burden over Skin Tone: Decomposing the Dermatology-AI Generalization Gap

The study examines why dermatology AI models, largely trained on light‑skinned, cancer‑focused images, perform poorly when applied to diverse patient populations. By comparing a cancer‑trained baseline, two dermatology foundation models, and a general‑purpose vision model on tone‑stratified and disease‑shifted datasets, the authors find that disease‑distribution shift, rather than skin‑tone underrepresentation, is the primary cause of generalization failure. Representation analysis shows that cancer‑specialized features lack transferable structure, while dermatology‑pretrained features maintain stronger clustering, and lightweight adaptation with about ten labeled examples per category can recover most performance.

By Nirajan Kunwor, Sanjaya Poudel, Quoc-Huy Trinh, Jahidul Arafat, Sunil Kumar Gaire
arXiv Machine Learning
Aug 20

MIFR: A Modality-Invariant and Fair Representation Framework for Skin Disease Classification

The paper introduces MIFR, a modality‑invariant and fair representation framework for skin disease classification that jointly processes clinical photographs and dermoscopic images using ViT‑based encoders. It employs a five‑component multi‑objective loss to balance classification accuracy, fairness across skin tones, class alignment, and modality invariance. Experiments on paired and external datasets demonstrate competitive predictive performance and fairness, with t‑SNE visualizations confirming alignment of embeddings from different modalities.

By Asonyu Senge Njih, Yvan Guifo Fodjo, Vianney Kengne Tchendji, Jerry Lacmou Zeutouo, Kerol Djoumessi
arXiv AI
5d ago

CG-HAF: An Interpretable Global-Local Lesion-Burden Fusion Framework for Ordinal Acne Severity Grading in Agentic Skincare Support

CG-HAF is a global‑local fusion framework for ordinal acne severity grading that explicitly combines holistic facial severity probabilities with structured lesion‑burden descriptors such as lesion count, detection confidence, and lesion area. The model uses a lightweight, interpretable classifier to produce the final grade, achieving statistically significant improvements over global‑evidence‑only baselines, especially for severe cases. Cross‑dataset testing reveals that strong performance within a dataset does not automatically transfer, largely due to mismatched grading criteria rather than detection failures.

By Muhammad Muhtasim Shahriar, Md. Naimur Asif Borno, Saad Aloteibi, Mohammad Ali Moni
arXiv Computer Vision
2d ago

Detail in Context: A Dual-Scale Machine Learning Framework for Mycosis Fungoides Detection

arXiv:2609.38560v1 Announce Type: new Abstract: Mycosis fungoides (MF) is a rare form of cutaneous T-cell lymphoma that is often misdiagnosed in early stages due to its visual similarity to benign in...

By Mohamed Hazem, Tarek Waleed, Omar Khaled, Nada Omar, Mahmoud Raslan, Marwa Mohamed Fawzy, Aya Fahim, Rania M. Mogawer, Ahmed Mourad, Kariman Mansour, Muhammad Rushdi
arXiv Computer Vision
Sep 11

A Comparative Evaluation of Pre-trained Convolutional Neural Networks for Melanoma Detection

The study compares five pre‑trained convolutional neural networks—ResNet50, VGG16, VGG19, MobileNet, and InceptionV3—for melanoma detection using dermatoscopic and histopathological image datasets. Accuracy varied across models and modalities, with ResNet50 achieving the highest scores (84% on HAM10000 and 83% on CR‑AI4SkIN) and InceptionV3 the lowest (71% on ISIC 2018). The results show that a model’s performance on dermatoscopic images does not necessarily predict its performance on histopathological images.

By Wagner Moreno Schmitz, Marco Antonio de Castro Barbosa, Thiago Magalh\~aes Amaral, Dalcimar Casanova, Jefferson Tales Oliva
arXiv AI
1d ago

LENS-GRF: Permutation-Invariant Lesion Evidence Network with Gated Residual Fusion for Acne Severity Grading and Multi-Rater Clinical Oracle Analysis

LENS‑GRF is a permutation‑invariant lesion evidence network that uses a Set‑Transformer and gated residual fusion to combine global facial context with localized lesion patches for four‑class acne severity grading. The framework integrates adaptive facial skin segmentation, a Vision Transformer prior, and a lesion set transformer that encodes spatial geometry, with a gating mechanism that modulates local residual contributions. In experiments on ACNE04 and PLSBRACNE01, the fully automated model achieved 80.82% accuracy, while using ground‑truth lesion annotations raised accuracy to 95.89% and a Quadratic Weighted Kappa of 0.9753; zero‑shot evaluation on the full cohort yielded 35.00% accuracy versus 42.50% for a global baseline, and oracle analyses on a 148‑subject cohort showed improved accuracy and QWK up to 47.97% and 0.5799.

By Muhammad Muhtasim Shahriar, M. F. Mridha