arXiv:2509.05259v2 Announce Type: replace
Abstract: Automatic Generation Control (AGC) plays a critical role in maintaining power balance across multi-area power systems. However, its complete relian...
By Ahmad Mohammad Saber, Alok Paranjape, Jehad Jilan, Niranjana Naveen Nambiar, Amr Youssef, Deepa Kundur
arXiv:2512.07419v3 Announce Type: replace
Abstract: Mixed-Precision Quantization (MPQ) liberates Deep Neural Networks (DNNs) from the Out-Of-Memory (OOM) bottleneck and has garnered increasing resear...
By Haidong Kang, Jun Du, Guo Yu
arXiv:2608.03854v4 Announce Type: replace
Abstract: Quantized large language models can run on consumer hardware, which motivates interest in on-premises processing of sensitive data. The reliability...
By Anton Rasmussen, Hong Qin
arXiv:2403.01805v2 Announce Type: replace-cross
Abstract: Shannon entropy regularization is widely adopted in optimal control due to its ability to promote exploration and enhance robustness, e.g., m...
By Yota Hashizume, Koshi Oishi, Shota Yamanaka, Kenji Kashima
arXiv:2609.24635v1 Announce Type: new
Abstract: When a language model reads an operation such as "Swap the contents of Box F and Box B", its forward pass writes keys and values for those tokens into...
By Lingfeng Wu, Behzad Shomali
arXiv:2609.24799v1 Announce Type: new
Abstract: Post-training quantization (PTQ) enables efficient deployment of large language models, and PTQ methods are usually optimized and evaluated with generi...
By Yeji Kim, Mi-Young Kim, Randy Goebel
arXiv:2609.24894v1 Announce Type: cross
Abstract: Whole-slide pathology images (WSIs) contain gigapixel-scale visual content, creating a major scalability challenge for slide-level multimodal large l...
By Ali Kerem Bozkurt, Baris Cem Bakay, Ibrahim Kulac, Cigdem Gunduz-Demir, Erkut Erdem, Aykut Erdem
arXiv:2609.24974v1 Announce Type: cross
Abstract: Agent harnesses, the external systems that mediate model-environment interaction, can substantially improve agent performance, but their gains remain...
By Haoran Ye, Yuxing Lu, Haonan Dong, Zhaochen Su, Guojie Song
arXiv:2508.03351v3 Announce Type: replace-cross
Abstract: Large language models (LLMs) have demonstrated remarkable capabilities across diverse language tasks, motivating their extension to vision-la...
By Yufei Xue, Yushi Huang, Lunjie Zhu, Jiawei Shao, Jun Zhang
arXiv:2609.22562v1 Announce Type: new
Abstract: Multi-vector visual document retrieval (VDR) models such as ColPali and ColNomic achieve strong accuracy by representing each document with hundreds to...
By Jianxin You, Kun Ni
arXiv:2609.23010v1 Announce Type: new
Abstract: Iterative text-to-motion generation delivers high-quality and semantically aligned motions but requires multiple network evaluations, resulting in subs...
By Hung Dinh, Binh Mai, Tran Quoc Bao Le, Lam Nguyen, Cong Tran
arXiv:2609.23153v1 Announce Type: new
Abstract: Video diffusion transformers are expensive because attention dominates long spatiotemporal token sequences. We identify the \emph{high-sparsity trap}:...
By Yuxi Liu, Haoyu Li, Zekun Zhang, Tengxu Sun, Yixiang Cai, Jiayong Li, Yifei Xia, Tianle Liu, Baole Ai, Ang Wang, Jiamang Wang, Lin Qu, Kai Zhang, Kun Yuan, Bin Cui
arXiv:2609.23268v1 Announce Type: new
Abstract: Blind image deconvolution (BID) is a prominent research topic in the field of imaging sciences, given its significant practical applications. Most exis...
By Qinghua Zhang, Xuesong Yang, Liangtian He, Liang-jian Deng, Jun Liu
arXiv:2609.23380v1 Announce Type: new
Abstract: Gaussian Splatting has enabled real-time novel view synthesis, but its tightly coupled geometry and appearance representation often require a large num...
By Zhiwei Li, Yijia Guo, Yishi Lu, Liwen Hu, Hong Rao, Shengbo Chen, Lei Ma
arXiv:2609.24526v1 Announce Type: new
Abstract: Physical AI requires models to ground visual and linguistic understanding in real-world environments while accounting for environmental constraints and...
By Foundation Model, Li Auto Inc
arXiv:2609.24825v1 Announce Type: new
Abstract: LiDAR point clouds acquired in underground environments exhibit severe geometric incompleteness due to occlusions and limited sensor viewpoints, making...
By Daisy Li, Kyle Gao, Quanyun Wu, Boris Jutzi, John S. Zelek, Jonathan Li
arXiv:2609.22277v1 Announce Type: cross
Abstract: Visual impairment affects over 2.2 billion people worldwide, yet conventional white canes cannot detect elevated hazards or provide semantic environm...
By Ali Akarma, Adeel Ahmad, Toqeer Ali Syed
arXiv:2603.28431v4 Announce Type: replace
Abstract: Although 3D Gaussian Splatting (3DGS) enables high-fidelity real-time rendering, its prohibitive storage overhead severely hinders practical deploy...
By Xuan Deng, Xiandong Meng, Hengyu Man, Qiang Zhu, Tiange Zhang, Debin Zhao, Xiaopeng Fan
Helix‑FNO is a teacher‑student framework that couples a 32‑state mechanistic model with a Fourier neural operator to learn the full solution operator for full‑scale treatment processes. The teacher generates a high‑fidelity dataset via Latin‑hypercube sampling and active learning, while the student learns in the spectral domain, enabling generalisation across varying influent profiles, controls, and plant layouts. The resulting operator achieves millisecond inference, three orders of magnitude faster than the mechanistic teacher, and is evaluated against physics‑informed and data‑driven surrogates on accuracy, dataset efficiency, and latency, positioning it on a speed‑accuracy Pareto front.
By Jiabao Zhao, Chuwei Wang, Jinxi Yang
The paper introduces OmicsBench, a new reasoning benchmark for multi‑omics sequences that includes 1,160 expert‑validated questions across DNA regulation, RNA processing, and protein function tasks, requiring traceable evidence chains. Evaluation of 17 large language models shows that scientific LLMs, while more accurate in classification, often lack valid evidence, suggesting shortcut learning. To address this, the authors propose tool‑augmented on‑policy distillation (TA‑OPD), a post‑training method that improves both evidence grounding and predictive performance across five Qwen3.5 models of varying sizes.
By Jie Ying, Zhefan Wang, Zihong Chen, Zhengqing Li, Jinzhe Li, Gang Li, Jian Liu, Fang Hu, Tao Luo, Zhonghang Yuan, Wanli Ouyang, Stan Z. Li, Fan Yang, Nanqing Dong