arXiv Machine Learning

Boosting Brain-to-Image Decoding with TRIBE v2 Data Augmentation

arXiv:2606. 06345v1 Announce Type: cross Abstract: Brain decoding is limited by the availability of labeled neural data, and remains challenging in low-data regimes.

arXiv Machine Learning
Sep 24

A Scaling Study for fMRI Foundation Models

The study investigates how data volume, model size, and training duration affect the performance of fMRI foundation models. Using over 200 datasets and 10,000 GPU‑hours, the authors find that larger models benefit more from additional data, and that at a fixed compute budget, increasing data yields greater gains than enlarging the model. By selecting optimal combinations of data, size, and duration, they produce models that outperform existing fMRI foundation models on out‑of‑distribution tasks while requiring less pretraining compute.

By Wenhao Ye, Xuanye Pan, Junfeng Xia, Junxiang Zhang, Mo Wang, Quanying Liu
arXiv Computer Vision
Sep 22

Brain-to-Image Generation: Reconstructing Visual Stimuli from EEG using Generative Adversarial Networks

The paper presents a reproducible single‑subject baseline for reconstructing visual stimuli from EEG using a temporal‑spatial convolutional encoder that maps averaged EEG signals to 512‑dimensional ViT-B/32 image features. On the THINGS‑EEG2 dataset, the model achieves 12.83%, 39.17%, and 58.00% image recall at ranks 1, 5, and 10, respectively, outperforming analytical chance levels. The study also shows that performance drops sharply when applying a model trained on one subject to others, and that direct conditional generators without external visual weights produce noise‑dominated outputs, indicating that only coarse semantic decoding is feasible under the tested protocol.

By Harshit Goyal
arXiv AI
Jul 28

Real-time Reconstruction of Human Visual Perception from fMRI

arXiv:2607. 22753v1 Announce Type: cross Abstract: Real-time closed-loop neurofeedback based on functional magnetic resonance imaging (fMRI) has led to important scientific and clinical advances.

By Rishab S. Iyer, Jiaxin Cindy Tu, Cesar Kadir Torrico Villanueva, Anish Mahishi, Ross P. Kempner, Jacob S. Prince, Ernest W. Lo, Akash Bhowmick, Hritik Arasu, Amaar Chughtai, Elizabeth A. McDevitt, Paul S. Scotti, Kenneth A. Norman
Hugging Face Trending Papers
Jun 3

Coarse-to-fine Hierarchical Architecture with Sequential Mamba for Brain Reconstruction

Understanding the relationship between deep visual representations and the human visual system is a fundamental challenge in computational neuroscience. While modern vision models achieve strong performance in image recognition, their correspondence with the hierarchical organization of the human visual cortex remains an open question.

arXiv AI
Sep 17

NeuroSketch: A Practical Design Recipe for Neural Decoding

NeuroSketch presents a practical design recipe for neural decoding, beginning with a comparative study of nine basic architectures that identifies CNN‑2D as the most effective. The recipe incorporates macro‑level gradual feature‑map expansion and early downsampling, along with micro‑level grouped convolutions, resulting in two variants—NeuroSketch‑Base (1.4M parameters) and NeuroSketch‑Large (4.2M parameters). Across nearly 5,000 experiments on eight tasks involving visual, auditory, and speech modalities and EEG, SEEG, and ECoG signals, both variants outperform ten baseline models on every task.

By Gaorui Zhang, Zhizhang Yuan, Jialan Yang, Junru Chen, Fanqi Shen, Li Meng, Yang Yang
arXiv AI
Sep 17

A Survey on Bridging EEG Signals and Generative AI: From Image and Text to Beyond

This survey reviews recent advances in converting non‑invasive EEG signals into images, text, and audio using generative AI techniques such as GANs, VAEs, transformers, and diffusion models. It summarizes datasets, feature‑encoding methods, evaluation metrics, and key challenges, noting that EEG‑to‑image models mainly use encoder‑decoder architectures, EEG‑to‑text leverages transformer language models, and EEG‑to‑audio maps signals to mel‑spectrograms for vocoder synthesis. The paper highlights the limitations of small, heterogeneous datasets, poor cross‑subject generalization, and the lack of standardized benchmarks, while providing open‑source resources to support reproducible research.

By Shreya Shukla, Jose Torres, Akshaj Murhekar, Christina Liu, Abhijit Mishra, Jacek Gwizdka, Shounak Roychowdhury