arXiv Machine Learning

A Common Measure of Communication for Speech Brain-Computer Interfaces

The paper introduces Open‑Vocabulary Mutual Information (OVMI), an information‑theoretic metric that quantifies how much of a user’s intended speech a speech brain‑computer interface (BCI) can convey relative to a reference word distribution. OVMI enables comparison of systems that use different vocabularies, recording methods, and datasets, revealing that conventional metrics like accuracy and word error rate can overstate performance. Using OVMI, the authors compare existing speech BCI systems, expose trade‑offs between vocabulary coverage and decoding accuracy, and show that optimizing vocabulary selection for OVMI can improve accuracy by up to 16.3% across three speech domains.

arXiv Computation and Language
1d ago

From Neurons to Conversation: Speech Brain-Computer Interfaces

arXiv:2609.36736v1 Announce Type: cross Abstract: Speech brain-computer interfaces (BCIs) aim to restore communication by transforming neural activity related to speech, language, or communicative in...

By Moein Khajehnejad, Forough Habibollahi, Tommaso Boccato, Margarida Sousa, Michal Olak, Francesco Jamal Sheiban, Matteo Ferrante
arXiv Machine Learning
Sep 4

The 2026 PNPL Competition: Word Classification and Efficient Cross-Subject Generalisation in LibriBrain100

The 2026 PNPL Competition builds on the 2025 PNPL effort by expanding the LibriBrain dataset to 32 new subjects and more within‑subject data, creating LibriBrain100. It introduces two tracks: a Deep track for high‑performance within‑subject word classification and a Broad track that tests cross‑subject generalisation with progressively less subject‑specific fine‑tuning data, down to 10 minutes. The competition aims to advance non‑invasive brain‑computer interfaces toward practical, clinically feasible communication restoration for people with profound paralysis.

By Francesco Mantegna, Gereon Elvers, Dulhan Jayalath, Gilad Landau, Tasha Kim, Miran \"Ozdogan, Luisa Kurth, Teyun Kwon, SungJun Cho, Benjamin Ballyk, Alex Fung, Anna Greer, Pratik Somaiya, Christian Herff, Yorguin Mantilla Ramos, Hamza Abdelhedi, Karim Jerbi, Greg Farquhar, Brendan Shillingford, Mark Woolrich, Oiwi Parker Jones
Hugging Face Trending Papers
Sep 3

The 2026 PNPL Competition: Word Classification and Efficient Cross-Subject Generalisation in LibriBrain100

The 2026 PNPL Competition builds on the 2025 PNPL effort by expanding the LibriBrain dataset to 32 new subjects and more within‑subject data, creating LibriBrain100. It introduces two tracks: a Deep track for high‑performance within‑subject word classification and a Broad track that tests cross‑subject generalisation with progressively less subject‑specific fine‑tuning data, down to about 10 minutes. The goal is to move toward a practical, non‑invasive brain‑computer interface that can restore communication for people with profound paralysis.

arXiv Machine Learning
Aug 27

LibriBrain100: One Hundred Hours of Broad and Deep MEG Data for Neural Speech Decoding at Scale

LibriBrain100 is a new large‑scale MEG dataset for speech decoding that contains over 100 hours of high‑quality recordings while subjects listened to naturalistic continuous speech. The dataset more than doubles the size of the original LibriBrain release, with a record 80 hours from a single subject and additional 40‑minute recordings from 32 subjects. The authors demonstrate the value of deep within‑subject data and broad multi‑subject data by achieving state‑of‑the‑art word‑classification performance and showing that supervised fine‑tuning can compensate for limited per‑subject data, all supported by open‑source tools and a public competition leaderboard.

By Francesco Mantegna, Dulhan Jayalath, Gereon Elvers, Tasha Kim, Benjamin Ballyk, Alex Fung, SungJun Cho, Teyun Kwon, Luisa Kurth, Miran \"Ozdogan, Gilad Landau, Pratik Somaiya, Natalie Voets, Mark Woolrich, Oiwi Parker Jones
arXiv Computation and Language
Sep 14

Not All Speech Is Intent: Adaptive Self-Correcting Inference Layer for Post-ASR False Wake-Up

The paper introduces ASCIL, a post‑ASR correction framework that re‑evaluates wake‑up intent by combining acoustic embeddings, linguistic cues, device context, and past misclassifications. ASCIL interprets both implicit (hesitation, disengagement, silence) and explicit (cancellation, repetition) signals as noisy indicators of misclassification, enabling online pattern updates without manual annotation. On a proprietary dataset of 3,667 interactions, ASCIL reduces errors by up to 54.27% relative on a session‑disjoint subset and 24.39% at a 0.90 threshold, while adding less than 60 ms of latency and improving intentional acceptance rates.

By Preeti Saraswat, Divya Neelagiri, Anil Yadav
arXiv Machine Learning
Aug 17

BCIJelly: An integrated ecosystem for brain-computer interface research

arXiv:2608. 13576v1 Announce Type: cross Abstract: Brain-computer interface (BCI) research relies on multistage computational pipelines, yet progress remains constrained by fragmented data formats, heterogeneous decoder implementations and hardware-specific deployment toolchains, and researchers lack an integrated workflow.

By Liyuan Han, Xinrui Yang, Tianyu Zheng, Qizhi Yang, Yitao Qin, Liang Chen, Qinglai Wei, Binjie Hong, Xinhe Zhang, Rui Xiong, Yong Gu, Mu-ming Poo, Bo Xu, Chengyu Li, Tielin Zhang
arXiv Machine Learning
Jul 22

Is EEG-to-Text Feasible in Real-World Scenarios? An In-Depth Analysis Using a Neuropsychology-Inspired Benchmark

arXiv:2607. 18749v1 Announce Type: new Abstract: Translating brain signals into text could restore communication for people with severe paralysis, yet practically usable systems to date rely on invasive electrocorticography (ECoG).

By Zihan Zhang (Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology), Yu Bao (Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology, Shanghai Innovation Institute), Xiao Ding (Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology), Tianyi Jiang (State Key Laboratory for Novel Software Technology, Nanjing University), Kai Xiong (Zhongguancun Laboratory)
arXiv Computation and Language
Sep 24

Brain-to-Language Decoding: Tasks, Signals, Methods, Evaluation, Practical Use and Beyond

The article surveys brain‑to‑language decoding, covering tasks, neural signals, methods, and evaluation across invasive and non‑invasive modalities. It traces the field’s evolution from constrained recognition to text generation, streaming speech, and facial animation, linking tasks to neural populations and decoder representations. The review highlights complementary decoding targets, shared representations, and the growing importance of calibration, feedback, and user control for online communication, while proposing a five‑level future trajectory toward bidirectional cognitive exchange.

By Yiqian Yang, Yiqun Duan, Chenyu Liu, Yiqi Wang, Xinliang Zhou, Chin-Teng Lin, Yu Zhang
arXiv AI
Aug 6

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

arXiv:2608. 04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instructions, signal processing, quantitative evidence, and scientific interpretation.

By Yangxuan Zhou, Sha Zhao, Yuning Chen, Chen Wu, Jiquan Wang, Shijian Li, Gang Pan