arXiv:2610.06902v1 Announce Type: new
Abstract: Retrieval-augmented generation grounds language models in external context, but for long documents flat top-$k$ retrieval can cluster on a single regio...
By Priyank Jayraj, Poonam Goyal, Navneet Goyal
arXiv:2610.06956v1 Announce Type: new
Abstract: Large speech language models have demonstrated strong capabilities in unified cross-modal understanding and generation, yet paralinguistic cues, especi...
By Jianan Pan, Yiwen Gu, Xinze Li, Rui Wang, Kejie Huang
arXiv:2610.07070v1 Announce Type: new
Abstract: Small language models are inexpensive to serve and can run on private infrastructure, but base models are often not good enough at multi-turn tool call...
By Aaron Fainman, Gabriela Kadlecov\'a, Maciej Gryka, Bartosz Kruszczy\'nski, Usman Zafar, C\'edric Archambeau, Aaron Klein, David Salinas, Selim Nowicki, Jacek Golebiowski
arXiv:2610.07109v1 Announce Type: new
Abstract: When an LLM judge scores an output, its score distribution retains uncertainty and disagreement information that is lost after scalar compression. We i...
By Yiqi Liu, Joseph James, Yang Wang, Kun Zhao, Chenghao Xiao, Chenghua Lin
arXiv:2610.07306v1 Announce Type: new
Abstract: Scientific search systems can find papers that are relevant to a question, but they generally do not assess the quality of the evidence that those pape...
By Matthew J. Vowels, Jamie Cummins
arXiv:2610.07365v1 Announce Type: new
Abstract: As LLMs increasingly assist scientific writing and peer review, detecting who wrote the text is no longer sufficient: we need to determine who contribu...
By Zhuoyang Zou, Abolfazl Ansari, Jiaxi Yang, Delvin Ce Zhang, Qian Chen, Dongwon Lee, Wenpeng Yin
arXiv:2610.07502v1 Announce Type: new
Abstract: Provider queries are clarifying requests sent by clinical documentation specialists to physicians to close gaps in the clinical note and ensure accurat...
By Joseph Paul Cohen, Raj Shah, Han-Chin Shing, Fang Wang, Susan Nguyen, Chaitanya Shivade, Jack Moriarty
arXiv:2410.21917v3 Announce Type: replace-cross
Abstract: The identifiability analysis of linear Ordinary Differential Equation (ODE) systems is a necessary prerequisite for making reliable causal in...
By Yuanyuan Wang, Biwei Huang, Wei Huang, Xi Geng, Mingming Gong
arXiv:2610.07722v1 Announce Type: new
Abstract: Activation steering provides a lightweight and flexible way to control large language model (LLM) behavior. However, effective steering requires more t...
By Haotian Yang, Huikang Jiang, Yucheng Wu, Wen-Jie Jiang, Chenpeng Wang, Yibin Lou, Liangming Pan
arXiv:2610.08037v1 Announce Type: new
Abstract: Language models frequently generate outputs in unintended languages or scripts, a phenomenon known as off-target generation. While existing research ha...
By David Kletz, Sandra Mitrovi\'c, Ljiljana Dolami\'c, Fabio Rinaldi
arXiv:2610.08085v1 Announce Type: new
Abstract: Speech-LLMs often exhibit prompt overfitting, where models solely trained on automatic speech recognition (ASR) instruction fail to generalize to new i...
By Hemant Yadav, Sunayana Sitaram, Roger Zimmermann, Rajiv Ratn Shah
arXiv:2610.08300v1 Announce Type: new
Abstract: Long-term conversational memory is becoming an integral component of modern LLM systems. Proposed architectures group records by topics and events, con...
By Michael Andreev
arXiv:2610.08585v1 Announce Type: new
Abstract: Large language models (LLMs) are increasingly relied upon to support ambient documentation and clinical reasoning. Here we examine the impact of a fail...
By Krithik Vishwanath, Brandon Ye, Anton Alyakin, John E. Markert, Aaron Hsieh, Micha{\l} Ma\'nkowski, Eric K. Oermann
arXiv:2610.08630v1 Announce Type: new
Abstract: Recently Large Language Models (LLMs) and LLM-based agents increasingly need to incorporate knowledge acquired after pretraining, e.g., domain facts, u...
By Haoyu Huang, Zhongwei Xie, Jiaxin Bai, Yisen Gao, Hong Ting Tsang, Wuganjing Song, Huihao Jing, Yufei Li, Yangqiu Song
arXiv:2610.08660v1 Announce Type: new
Abstract: Background: Biomedical AI can generate plausible explanations without reliably verifying whether each statement is supported by patient-specific eviden...
By Mariya Miteva, Maria Nisheva-Pavlova
arXiv:2610.08703v1 Announce Type: new
Abstract: In K-12 mathematics tutoring, student-tutor dialogue provides rich evidence of learners' problem-solving processes and sources of difficulty. Learning...
By Clayton Cohn, Joyce Fonteles, Kirk Vanacore, Gianni Mazza, Candida Crawford, Tom Hooper, Gautam Biswas, Rene Kizilcec
arXiv:2610.08747v1 Announce Type: new
Abstract: A common approach to measuring bias in Large Language Models is to compare the log-likelihoods of two contrastive stereotype sentences. We argue that s...
By Nataliya Stepanova, Ivan Titov, Emily Allaway, Bj\"orn Ross
arXiv:2610.08063v1 Announce Type: cross
Abstract: This paper presents the HINTT system submitted to the 2nd Challenge and Workshop on Multilingual Conversational Speech Language Model (MLC-SLM). We a...
By Takanori Ashihara, Kohei Matsuura, Masato Mimura
arXiv:2510.22954v2 Announce Type: replace
Abstract: Language models (LMs) often struggle to generate diverse, human-like creative content, raising concerns about the long-term homogenization of human...
By Liwei Jiang, Yuanjun Chai, Margaret Li, Mickel Liu, Raymond Fok, Nouha Dziri, Yulia Tsvetkov, Maarten Sap, Alon Albalak, Yejin Choi
arXiv:2509.26189v2 Announce Type: replace
Abstract: The rapid proliferation of Large Language Models has intensified the challenge of distinguishing LLM-generated text from human writing in non-Engli...
By Trieu Hai Nguyen, Sivaswamy Akilesh