arXiv AI

CytoCLIP: Learning Cytoarchitectural Characteristics in Developing Human Brain Using Contrastive Language Image Pre-Training

arXiv:2601. 12282v2 Announce Type: replace-cross Abstract: The functions of different regions of the human brain are closely linked to their distinct cytoarchitecture, which is defined by the spatial arrangement and morphology of the cells.

arXiv Machine Learning
Aug 12

Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy Datasets

arXiv:2608. 10657v1 Announce Type: cross Abstract: Leukemia cell image classification is challenged by real-world domain shifts from acquisition, staining, illumination, and site protocols, causing single-dataset models to generalize poorly in real clinical scenarios.

By Carlos Zamora, Hiram Zuniga, Ulises Orozco-Rosas, Kenia Picos
Hugging Face Trending Papers
Aug 18

TEAMS: Text-prompted spatiotEmporal dual-heAd Mamba Snake

The paper introduces TEAMS, a vision‑language Mamba snake framework that enhances deep snake instance segmentation. It adds a Spatiotemporal Snake Evolution Strategy to handle complex shapes, a Contour Morphology‑Aware Mamba to improve fine‑grained detail capture, and a Text‑prompted Collaborative Dual‑Head Snake to integrate textual cues and reduce detection errors. Experiments on five medical imaging datasets show TEAMS surpasses existing methods, achieving significant gains in mDice and mBF metrics.

Hugging Face Trending Papers
Jun 3

Coarse-to-fine Hierarchical Architecture with Sequential Mamba for Brain Reconstruction

Understanding the relationship between deep visual representations and the human visual system is a fundamental challenge in computational neuroscience. While modern vision models achieve strong performance in image recognition, their correspondence with the hierarchical organization of the human visual cortex remains an open question.

arXiv AI
Jun 11

Brain-IT-VQA: From Brain Signals to Answers

arXiv:2605. 29588v2 Announce Type: replace-cross Abstract: Decoding visual content from fMRI signals recorded while a person views images, and specifically answering questions about the seen images, is a long-standing challenge.

By Roman Beliy, Matias Cosarinsky, Oliver Heinimann, Navve Wasserman, Michal Irani
arXiv AI
Jul 17

Parameter-efficient Prompt Tuning of Vision Foundation Model With Adaptive Focal Loss for Interpretable MCI Screening

arXiv:2607. 15047v1 Announce Type: cross Abstract: Mild Cognitive Impairment is a critical early stage of cognitive decline that frequently precedes Alzheimer's disease, yet its automated detection from neuropsychological drawing tests remains fundamentally constrained by data scarcity, class imbalance, and diagnostic ambiguity near clinical boundaries.

By Javad Khoramdel, Farhad Hoseyni, Amirhossein Nikoofard