arXiv Computer Vision By Shubham Gajjar, Deep Joshi, Avi Poptani, Vishal Barot

VGG16-MCA UNet: Whole-Tumor Segmentation in 2D FLAIR MRI with Decoder-Side Channel Attention

Read the original on arXiv Computer Vision →

VGG16-MCA UNet is a hybrid neural network that combines an ImageNet‑pretrained VGG16 encoder with a decoder enhanced by a Multi‑Channel Attention module, trained using Focal Tversky loss to address class imbalance. The model was evaluated as a 2‑D, FLAIR‑only whole‑tumor segmenter on BraTS 2020 and LGG datasets, achieving a pixel‑level Dice of 95.10 % on BraTS and 88.32 % on LGG in a 5‑fold cross‑validation setting. Inference time is 66.32 ms per 256×256 slice on a single RTX 2060, only slightly slower than a VGG16‑UNet without attention. whyItMatters":"The study provides a reproducible 2‑D FLAIR baseline for whole‑tumor segmentation, demonstrating high Dice scores and detailed reporting of training and evaluation protocols."

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv AI
Sep 1

Subtraction-Based Tumor Segmentation and Lesion-Centered pCR Prediction for the MAMA-MIA Challenge

Team FME submitted a method for the MAMA-MIA Challenge that tackles primary tumor segmentation and pathological complete response (pCR) prediction using dynamic contrast‑enhanced breast MRI. For segmentation, they employed a five‑fold residual‑encoder nnU‑Net ensemble trained on the first post‑contrast minus pre‑contrast image, augmented with mirroring test‑time augmentation and largest‑connected‑component filtering, achieving a Dice score of 0.713 and a normalized Hausdorff distance of 0.099. For pCR prediction, they ensembled 25 pretrained 3D video classifiers on lesion‑centred crops from the pre‑contrast and first two post‑contrast volumes, reaching a balanced accuracy of 0.541 and an equalized‑odds disparity of 0.212, and ranked second in both tasks.

By Kai Geissler, Raphael Sch\"afer
arXiv Machine Learning
Sep 16

NeuroTS-Net: Multi-Class Semantic Segmentation of Pediatric Brain Tumors in Multi-Modal MRI

NeuroTS-Net is a 3‑D encoder‑decoder CNN designed for multi‑class semantic segmentation of pediatric brain tumors in multi‑modal MRI. It uses a dual‑scale raw‑detail stream, adaptive low‑resolution context selection, and detail‑preserving multipath downsampling to maintain fine intensity and boundary information while modeling broader tumor context. Trained on the BraTS 2026 pediatric dataset, it outperformed nnU‑Net and MedNeXt, achieving Dice scores of 0.938/0.937 on internal validation and 0.927/0.926 on the official challenge set.

By Darius Peteleaza, Razvan-Gabriel Dumitru, Bogdan Neamtu, Arpad Gellert, Mariana Sandu, Claudiu Matei
arXiv Computer Vision
Aug 27

Synergistic Modality-and-Slice Memory Framework for Cross-Modal 3D Brain Tumor Segmentation

The paper introduces MSM‑Seg, a dual‑memory segmentation framework for 3D multi‑modal brain tumor segmentation. It combines a modality‑and‑slice memory attention module to capture cross‑modal and spatial‑slice dependencies, a multi‑scale category‑agnostic prompt encoder for whole‑tumor guidance, and a modality‑adaptive fusion decoder to integrate complementary decoding information. Experiments on various MRI datasets show that MSM‑Seg surpasses state‑of‑the‑art methods for metastases and glioma tumor segmentation.

By Yuxiang Luo, Qing Xu, Hai Huang, Yuqi Ouyang, Xiangjian He, Zhen Chen, Wenting Duan, Jiebo Luo