arXiv Machine Learning By Benjamin Ballyk, Teyun Kwon, Miran \"Ozdogan, Oiwi Parker Jones

Physiological Noise Augmentation Improves Non-Invasive Brain-to-Speech

Read the original on arXiv Machine Learning →

arXiv:2607. 05165v1 Announce Type: new Abstract: Non-invasive brain-to-speech decoding aims to restore communication to patients suffering from neurodegenerative disease, without the risks of neurosurgery.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Aug 20

Accurate Decoding of Natural Sentences from Non-Invasive Brain Recordings

Brain2Qwerty v2 is a model that decodes natural sentences from real‑time magnetoencephalography (MEG) recordings, achieving an average word error rate of 39% across 22,000 sentences typed by nine subjects. The model uses character, word, and sentence‑level representations and shows that decoding accuracy improves log‑linearly with more data, narrowing the gap to intracranial brain‑computer interfaces. AI contributes by replacing hand‑crafted event detection with deep learning, fine‑tuning large language models for semantic extraction, and employing AI agents to refine the decoding pipeline through automated code development.

By Mingfang Zhang, Jarod L\'evy, Cedric Rommel, J\'er\'emy Rapin, Corentin Bel, Julie Bonnaire, Daniel Nieto, Pierre Bourdillon, Svetlana Pinet, St\'ephane d'Ascoli, Thomas Moreau, Jean-R\'emi King
arXiv AI
Aug 25

Cross-Subject Generalization in Decoding Perceived Speech from Non-Invasive Brain Recordings

The paper introduces a Cross-Subject Perceived Speech Decoding (CPSD) framework that tackles the challenge of decoding perceived speech from non‑invasive brain recordings across different subjects. CPSD uses a two‑stage training process: first, contrastive learning pre‑trains a source model on multiple subjects to capture shared representations; second, personal specialization fine‑tunes the model for a target subject by extracting consistent components and further training on that subject’s data. A Positional Encoding‑based Spatial Attention (PESA) module is added to remap MEG/EEG data into a standardized reference space, improving cross‑subject consistency. Evaluations on three datasets (Armeni 2022, PKUEEG 2025, Broderick 2018) show that CPSD outperforms baseline methods by more than 6.8%, 15.4%, and 15.8% in Top‑10 accuracy, demonstrating its effectiveness, efficiency, and robustness.

By Aoke Zhang, Bo Wang, Xihong Wu, Heping Cheng, Jing Chen
arXiv Computation and Language
4d ago

From Neurons to Conversation: Speech Brain-Computer Interfaces

arXiv:2609.36736v1 Announce Type: cross Abstract: Speech brain-computer interfaces (BCIs) aim to restore communication by transforming neural activity related to speech, language, or communicative in...

By Moein Khajehnejad, Forough Habibollahi, Tommaso Boccato, Margarida Sousa, Michal Olak, Francesco Jamal Sheiban, Matteo Ferrante