arXiv:2610.06892v1 Announce Type: cross
Abstract: Learning from human preference data is the dominant route to aligning language models with human values. In linear social choice, where rewards are l...
By Soumya Nasipuri, Sayak Ray Chowdhury, Sanjukta Roy
arXiv:2610.06927v1 Announce Type: cross
Abstract: The key-value (KV) cache of autoregressive transformers grows linearly with context length and dominates memory at long context. Most training-free r...
By Sara Abdali, Jongwoo Ko, Pashmina Cameron
arXiv:2610.06977v1 Announce Type: cross
Abstract: Multimodal large language models (MLLMs) remain vulnerable to transferable adversarial examples, especially in black-box settings where only open-sou...
By Xiaojun Jia, Simeng Qin, Yiming Li, Jie Liao, Sensen Gao, Ke Ma, Yang Liu, Xiaochun Cao
arXiv:2610.06993v1 Announce Type: cross
Abstract: Evolution Strategies (ES) enable memory efficient full parameter fine-tuning of large language models (LLMs) using only forward computation. However,...
By Zhishen Sun, Hongzhan Wang, Sizhe Dang, Guang Dai, Haishan Ye
arXiv:2610.06996v1 Announce Type: cross
Abstract: Block diffusion language models keep a large key-value (KV) cache throughout generation and attend to it at every denoising step, limiting both memor...
By Gleb Molodtsov, Ekaterina Alimaskina, Evgeny Uskov, Artur Zagitov, Aleksandr Beznosikov
arXiv:2610.08647v1 Announce Type: new
Abstract: LLM-based agents solve complex multi-step tasks, but sequential execution incurs substantial latency. In principle, parallelizing work across multiple...
By Yexiong Lin, Shanshan Ye, Yu Yao, Zhen Fang, Bo Han, Tongliang Liu
arXiv:2610.07098v1 Announce Type: cross
Abstract: Large transformer-based models increasingly depend on multi-GPU execution, which requires frequent collective communication among GPUs. Existing comm...
By Keyvan Dadashzadeh, Yuehong Zhou, Minyu Cui, Miquel Pericas
arXiv:2610.07115v1 Announce Type: cross
Abstract: The order in which candidate responses are presented can change an LLM judge's verdict. Detecting such a position flip ordinarily requires judging ea...
By Hashmath Shaik, Gnaneswar Villuri, Alex Doboli
arXiv:2610.07132v1 Announce Type: cross
Abstract: Croissant has emerged as a standard for machine-readable dataset metadata, yet populating its fields remains labor-intensive and requires careful rea...
By Berke Arda, Ahmetcan Yavuz, Paul Gerry, Sebastian Lobentanzer, Nobin Sarwar, Joan Giner-Miguelez, Kongtao Chen, Luyao Zhang, Mrinmaya Sachan, Mubashara Akhtar
arXiv:2610.07207v1 Announce Type: cross
Abstract: Mixture-of-Experts (MoE) transformers scale capacity by activating only a few experts per token, but this sparsity creates a hidden reliability probl...
By Xin Teng, Muxiao Li, Hongyi Wen
arXiv:2610.07224v1 Announce Type: cross
Abstract: Clinical notes capture most of what is documented about a patient's care, but they cannot be used for research until protected health information (PH...
By Jose D. Posada, Somalee Datta, Priya Desai
arXiv:2610.07276v1 Announce Type: cross
Abstract: Deployment-time safety of language models is commonly implemented through runtime guardrails such as input moderation, routing, retrieval verificatio...
By Xingru Zhou, Luis Sentis, Aarti Choudhary
arXiv:2610.07298v1 Announce Type: cross
Abstract: Cyber threat analysis increasingly depends on evidence distributed across vendor advisories, vulnerability databases, and threat intelligence sources...
By Luoxi Tang, Yuqiao Meng, Ankita Patra, Weicheng Ma, Muchao Ye, Zhaohan Xi
arXiv:2610.07335v1 Announce Type: cross
Abstract: Improving the reliability of large language model (LLM) agents in long-horizon decision-making remains a key challenge. When deployed as autonomous a...
By Heewon Park, Somin Im, Minhae Kwon
arXiv:2610.07339v1 Announce Type: cross
Abstract: Tactical Combat Casualty Care (TC3) requires responders to connect visual observations of injuries and interventions with established clinical guidan...
By Junseob Kim, Jade Chng, Ayman Ali, Victor Moas, Yichun Lee, Po-Chun Chin, Sunil Hwang, Rishikesan Kamaleswaran
arXiv:2610.07460v1 Announce Type: cross
Abstract: Inserting objects into existing 3D scenes requires more than selecting a plausible location:
the inserted object must also fit local geometry while...
By Tzu-Hsin Hsieh, Ricardo Marroquim
arXiv:2610.07532v1 Announce Type: cross
Abstract: LLMs have advanced rapidly, raising growing concerns about their safety. Recent work has proposed approaches to detect and defend against attacks inc...
By Wonjun Lee, Kyungsik Yang, Gaeun Ji, Vaidehi Patil, Haon Park, Bumsub Ham, Mohit Bansal, Suhyun Kim
arXiv:2610.07654v1 Announce Type: cross
Abstract: On-policy distillation (OPD) has attracted growing attention as an effective way to transfer capabilities from teacher models to student models. Rece...
By Jian Luo, Kehan Qi, Qingqiao Hu, Meilong Xu, Jiacheng Qiu, Weimin Lyu, Jiawei Zhou, Chao Chen
arXiv:2610.07676v1 Announce Type: cross
Abstract: Research on transformer expressivity shows whether a transformer is capable of solving a given task, but gives little indication of whether the solut...
By Yijia Jessica Zhu, David Chiang
arXiv:2610.07742v1 Announce Type: cross
Abstract: Optimized kernels such as FlashAttention and FlashDecoding are crucial for accelerating today's large models. Most of them are handwritten by experts...
By David Pissarra, Jinkun Lin, Haitian Jiang, Aurojit Panda, Jinyang Li