Lumen is a pathology vision‑language model that aligns frozen unimodal foundation models (Virchow2 and BioMedBERT) using rank‑4 adapters and projection heads, training only 0.40% of the total parameters on the QUILT‑1M corpus. It achieves the highest mean chance‑corrected balanced accuracy (0.546) across nine zero‑shot patch benchmarks and demonstrates strong performance on lymph‑node metastasis detection, with AUROC scores of 0.964 internally and 0.955 externally. While it ranks third in cross‑modal retrieval, Lumen’s low‑parameter training yields competitive results at both patch and slide levels.
By Kiarash Tajbakhsh, Abdelrahman Faqieh, Michael Jopiti, Javier Garcia-Baroja, Philipp Zens, Branislav Zagrapan, Yuri Tolkach, Martin D. Berger, Aurel Perren, Bastian Dislich, Inti Zlobec, Amjad Khan
arXiv:2607. 03581v1 Announce Type: cross Abstract: The widespread adoption of facial masks, accelerated by COVID-19 and mandated in security-sensitive settings, has exposed limitations of conventional face recognition systems.
By Dana A Abdullah
arXiv:2609.13237v1 Announce Type: cross
Abstract: Orthodontic report generation from intraoral data is normally cast as multimodal captioning, yet the released Bite2Text scan pairs are supplied alrea...
By Ajo Babu George, Govind Arun, Sidharth N Krishna, Uma Ranjan
arXiv:2512.15774v5 Announce Type: replace
Abstract: The absence of large-scale masked face datasets challenges masked face detection and recognition. We propose a two-step generative data augmentatio...
By Yan Yang, George Bebis, Mircea Nicolescu
arXiv:2609.15144v1 Announce Type: new
Abstract: Purpose: Paired pre- and post-operative photographs are the standard unit of evidence for plastic surgical outcomes, yet no objective metric verifies w...
By Derrick Lin, Samantha Rabinovich, Joclin Rabinovich, Kassra Garoosi, Sumun Khetpal, Evan Delanoy, Neel Bhardwaj, Jason Roostaeian
arXiv:2607. 21318v1 Announce Type: cross Abstract: Replacing an object with one that differs in category or shape requires complete source removal, natural target formation unconstrained by the source silhouette, and preservation of unrelated content.
By Jian Zhang, Zhijun Zhang