arXiv AI By Negin Kafee Hernashki, Soumick Chatterjee

MIRTO: a registration-gated, multiverse-tested evaluation protocol for unsupervised anomaly segmentation in brain MRI

Read the original on arXiv AI →

MIRTO is an evaluation protocol for unsupervised anomaly segmentation in brain MRI that explicitly documents key methodological choices—such as registration alignment, threshold setting, and false‑positive budgeting—and measures their impact. It applies a registration check, uses validation data for thresholding, reports realized false‑positive volumes, and repeats each comparison across 15,552 evaluation pipelines with bootstrap intervals. In a study on four UAD methods and 312 BraTS 2020 subjects, MIRTO revealed that an axis‑order mismatch dramatically lowered a diffusion model’s voxel AUROC, and that many performance differences were driven by lesion definition and threshold transfer rather than model quality.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 15

Catching magnetic resonance imaging outliers in artificial intelligence-supported radiotherapy workflows: unsupervised detection and localization of image anomalies using deep learning

arXiv:2605. 24609v2 Announce Type: replace-cross Abstract: Artificial intelligence is increasingly integrated into radiotherapy workflows, yet such pipelines remain vulnerable to out-of-distribution image data that may introduce unexpected behavior in clinical tasks.

By Mustafa Kadhim, Viktor Rogowski, Emilia Persson, Camila Gonzalez, Andr\'e Haraldsson, Sofie Ceberg, Mikael Nilsson, Malin K\"ugele, Sven B\"ack, Christian Jamtheim Gustafsson
arXiv Computer Vision
Sep 11

Pre- and Post-Treatment Brain Metastases Segmentation Using nnU-Net with Post-Processing for BraTS 2026

The paper presents a segmentation pipeline for brain metastases in both pre‑ and post‑treatment cases using a 5‑fold nnU‑Net ResEnc‑L ensemble trained on 1,296 four‑modality cases. A rule‑based post‑processing cascade improves the lesion‑wise Dice similarity coefficient (LW‑DSC) for enhancing tumour, tumour core, whole tumour, and resection cavity sub‑regions, achieving LW‑DSC scores of 0.733, 0.751, 0.713, and 0.549 respectively on the official validation leaderboard. The authors conduct a five‑fold out‑of‑fold analysis to validate the robustness of each post‑processing stage, provide a mechanistic explanation of LW‑DSC behaviour, and report thirteen negative results that challenge common intuitions, with all code released under Apache‑2.0.

By Haobin Liu, Xin Wang