EndoFSA is a GAN-based model designed for endoscopic few-shot image generation, addressing the scarcity of pathological samples in wireless capsule endoscopy (WCE) data. It adapts a generator pretrained on abundant normal images to abnormal domains by updating only a small set of rank-constrained modulation parameters while keeping the rest of the weights frozen, thereby preserving anatomical priors and preventing mode collapse. The method incorporates perceptual boundary regularization and cluster-wise diversity control, operates without pixel-level annotations, and demonstrates that synthetic abnormal images can match real images in downstream classification performance.
By Panagiota Gatoula, Grigoris Karypidis, Dimitris K. Iakovidis
arXiv:2310. 07895v2 Announce Type: replace Abstract: This paper presents a method to efficiently classify the gastroenterologic section of images derived from Video Capsule Endoscopy (VCE) studies by exploring the combination of a Convolutional Neural Network (CNN) for classification with the time-series analysis properties of a Hidden Markov Model (HMM).
By Julia Werner, Christoph Gerum, Moritz Reiber, J\"org Nick, Oliver Bringmann
arXiv:2505. 11034v2 Announce Type: replace-cross Abstract: Robust machine learning depends on clean data, yet current image data cleaning benchmarks rely on synthetic noise or narrow human studies, limiting comparison and real-world relevance.
By Fabian Gr\"oger, Simone Lionetti, Philippe Gottfrois, Alvaro Gonzalez-Jimenez, Ludovic Amruthalingam, Elisabeth Victoria Goessinger, Hanna Lindemann, Marie Bargiela, Marie Hofbauer, Omar Badri, Philipp Tschandl, Arash Koochek, Matthew Groh, Alexander A. Navarini, Marc Pouly
arXiv:2608.24281v1 Announce Type: new
Abstract: Reducing annotation requirements remains a key challenge in developing robust medical object detectors. To address this, Vision-Language (VL) object de...
By Sheethal Bhat, Bogdan Georgescu, Awais Mansoor, Mathias Zinnen, Pranjal Sahu, Florin C. Ghesu, Sasa Grbic, Andreas Maier
arXiv:2610.00414v1 Announce Type: new
Abstract: Foundation models pretrained on large-scale datasets demonstrate strong transferability to medical imaging tasks. However, understanding how their late...
By Michael D. Vasilakakis (Department of Computer Science and Biomedical Informatics, University of Thessaly, Lamia, Greece), Dimitris K. Iakovidis (Department of Computer Science and Biomedical Informatics, University of Thessaly, Lamia, Greece)
arXiv:2606. 16868v1 Announce Type: cross Abstract: While federated learning (FL) enables collaborative medical image segmentation without centralizing sensitive data, real-world deployment is frequently complicated by cross-site label imperfections such as contour disagreement, missing or additional structures, and confused labels.
By Markus Bujotzek, Dimitrios Bounias, Stefan Denner, Ralf Floca, Maximilian Fischer, Peter Neher, Klaus Maier-Hein