arXiv Computer Vision

Research on Deep Learning-Based Semantic Segmentation Algorithms for Subcortical Brain Structures

arXiv Computer Vision
Aug 27

Deep Learning Segmentation of Diffusion-Weighted MRI Acute Ischaemic Stroke: A Pragmatic Evaluation Across Three Datasets

The study evaluated a pragmatic deep‑learning approach for segmenting acute ischemic stroke lesions on diffusion‑weighted MRI. Using a self‑configured nnU‑Net trained on 1,744 cases and tested on 436, the baseline model achieved a median Dice similarity coefficient of 0.84, outperforming the DeepISLES ensemble, especially for smaller infarcts. The approach required minimal preprocessing and fast inference, suggesting it could streamline clinical stroke imaging workflows.

By Atle Bj{\o}rnerud, Till Schellhorn, Thor H. Skatt{\o}r, Terje Nome, Jon Andr\'e Ottesen, Anne Hege Aamodt, Bradley J MacIntosh
arXiv Machine Learning
Sep 16

NeuroTS-Net: Multi-Class Semantic Segmentation of Pediatric Brain Tumors in Multi-Modal MRI

NeuroTS-Net is a 3‑D encoder‑decoder CNN designed for multi‑class semantic segmentation of pediatric brain tumors in multi‑modal MRI. It uses a dual‑scale raw‑detail stream, adaptive low‑resolution context selection, and detail‑preserving multipath downsampling to maintain fine intensity and boundary information while modeling broader tumor context. Trained on the BraTS 2026 pediatric dataset, it outperformed nnU‑Net and MedNeXt, achieving Dice scores of 0.938/0.937 on internal validation and 0.927/0.926 on the official challenge set.

By Darius Peteleaza, Razvan-Gabriel Dumitru, Bogdan Neamtu, Arpad Gellert, Mariana Sandu, Claudiu Matei
arXiv AI
Aug 28

Pixel Wised Lesion Prediction on COVID-19 CT Imagery: A Comparative Analysis of Automated Image Segmentation Architectures

The study evaluates four deep‑learning segmentation architectures—Unet, PSPNet, Linknet, and FPN—paired with six pre‑trained encoders to predict COVID‑19 lesions in CT images. Experiments on three COVID‑19 CT datasets show high accuracy, achieving a maximum binary F1‑score of 98% and multi‑class F1‑scores of 75% and 77%. The work aims to provide a standardized performance benchmark for medical image segmentation and a reference for other imaging scenarios.

By Sarmad Khan, Basim Azam, Arslan Shaukat
arXiv Computer Vision
Aug 31

3D MRI-Based Alzheimer's Disease Classification Using Multi-Modal 3D CNN with Leakage-Aware Subject-Level Evaluation

The paper presents a multimodal 3D convolutional neural network that classifies Alzheimer’s disease using raw OASIS 1 MRI volumes. It fuses structural T1 images with gray matter, white matter, and cerebrospinal fluid probability maps to capture complementary neuroanatomical information. Evaluated with 5‑fold subject‑level cross‑validation, the model achieves a mean accuracy of 72.34 % and an ROC AUC of 0.7781, with GradCAM visualizations highlighting anatomically relevant regions such as the medial temporal lobe and ventricles.

By Md Sifat, Sania Akter, Akif Islam, Md. Ekramul Hamid, Abu Saleh Musa Miah, Najmul Hassan, Md Abdur Rahim, Jungpil Shin
arXiv Computer Vision
Sep 4

Explainable Convolutional Neural Networks for Retinal Fundus Classification and Cutting-Edge Segmentation Models for Retinal Blood Vessels from Fundus Images

The paper presents a two‑pipeline framework for retinal fundus analysis that combines four‑class disease classification with vessel segmentation. It fine‑tunes eight ImageNet‑pretrained CNNs on the FIVES dataset, applies five gradient‑based explanation methods to assess model interpretability, and benchmarks ten U‑Net variants—including transformer‑based and attention‑enhanced architectures—on the FIVES and DRIVE datasets. The best classification results come from ResNet101 (94.17% accuracy), while the strongest segmentation performance is achieved by Attention U‑Net with a ResNet101V2 backbone, improving DRIVE IoU from 60.80% to 64.83%.

By Fatema Tuj Johora Faria, Mukaffi Bin Moin, Pronay Debnath, Asif Iftekher Fahim, Faisal Muhammad Shah
Hugging Face Trending Papers
Jun 27

A Deep Multiscale Neural Network for Accurate Neurological Disorder Detection from MRI Scans and Real-Time Web Deployment

Neurological disorders involve diverse pathologies of the brain and nervous system, making early and accurate detection essential. While many deep CNNs have been developed for MRI-based classification of neurological disorders, most are optimized for binary tasks and often fail to capture the multi-class features needed to distinguish subtle anatomical differences across conditions.

arXiv Computer Vision
Sep 25

Lightweight Vision Transformer-Based U-Net for Brain Tumor Segmentation from MRI

The paper introduces a lightweight Vision Transformer‑based U‑Net for brain tumor segmentation from MRI, combining U‑Net’s hierarchical feature extraction with a compact ViT bottleneck to capture both local and global context. With only 2.6 million trainable parameters, the model achieves a mean Intersection over Union of 0.8100 and a Dice score of 0.8446 on the TCGA LGG dataset, surpassing the baseline U‑Net by 3.75% and 3.15% respectively. Extensive quantitative and qualitative analyses, including confusion matrices, precision‑recall curves, and tumor size dependency studies, demonstrate the method’s effectiveness and robustness.

By Sheekar Banerjee, Md. Srabon Chowdhury, Md. Mahbub Hasan Akash, Ishtiak Al Mamoon