arXiv AI By Xinxin Li, Huiyao Chen, Meishan Zhang, Yunxin Li, Zulong Chen, Zhibo Ren, Xiaoqing Dong Baotian Hu, Min Zhang

Ontology Memory-Augmented ASR Correction for Long Text-Speech Interleaved Conversations

Read the original on arXiv AI →

arXiv:2606. 13464v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) correction has traditionally focused on isolated utterances or short local contexts.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.

Hugging Face Trending Papers
Jun 2

Efficient ASR Training with Conversations that Never Happened

Conversational ASR for lower-resource languages and niche domains is limited by the scarcity of domain-matched multi-speaker training data. We propose an augmentation pipeline that generates scenario-level dialogues with participant metadata, maps speaker attributes to TTS voice profiles, and assembles synthesized utterances into speaker-aware simulated conversations.