arXiv:2609.24357v1 Announce Type: new
Abstract: Cross-domain Named Entity Recognition (CD-NER) aims to transfer the rich knowledge in the source domain to the target domain. Recent studies adopting d...
By Jingyu Wang, Shijie Wu, Fusheng Jin
arXiv:2410.17021v2 Announce Type: replace
Abstract: Large Language Models with chain-of-thought prompting, such as OpenAI-o1, have shown impressive capabilities in natural language inference tasks. H...
By Xiaochen Wang, Liang Chen, Reza Haf Zhe Yang, Yiru Wang, Xiangdi Meng, Kunhao Pan, Zhifang Sui, Junqing He
arXiv:2604.08974v2 Announce Type: replace
Abstract: Uncertainty quantification techniques measure confidence in language model outputs to support critical applications like hallucination detection an...
By Lorenzo Jaime Yu Flores, Cesare Spinoso di-Piano, Jackie Chi Kit Cheung
arXiv:2606.22419v3 Announce Type: replace
Abstract: A recent Nature Medicine study reports that general-purpose frontier LLMs outperform specialized retrieval-augmented clinical tools on medical benc...
By Madhulatha Mandarapu, Sandeep Kunkunuru
arXiv:2607.01733v2 Announce Type: replace
Abstract: Speech-LLM integration has shown promising results by leveraging extensive textual pretraining, yet its specific benefits for automatic speech reco...
By Ruchao Fan, Yiming Wang, Rui Zhao, Liliang Ren, Keqi Deng, Xiaoyang Chen, Ali Zare, Bo Ren, Yuxuan Hu, Junkun Chen, Yan Huang, Yelong Shen, Jinyu Li
arXiv:2608.04186v3 Announce Type: replace
Abstract: This paper presents a conceptual framework for developing an electronic explanatory dictionary of the Tajik language using large language models (L...
By Mullosharaf K. Arabov, Saidali M. Pirzoda, Behruz A. Sultonov
arXiv:2608.24327v2 Announce Type: replace
Abstract: With the advent of Large Language Models and its instruction following capabilities a promising application is the task of summarization. Within th...
By Enes Yavuz Ugan, Fabian Retkowski, Yuka Ko, Thai-Binh Nguyen, Maike Z\"ufle, Jan Niehues, Alexander Waibel
arXiv:2609.24075v1 Announce Type: new
Abstract: Computer-vision systems used for construction monitoring can degrade under adverse environmental and visual conditions, yet such conditions remain unde...
By Viet Huy Duong, Ruoxin Xiong, Md Abdullah Al Forhad, Weishi Shi
arXiv:2511.18921v2 Announce Type: replace
Abstract: Backdoor attacks undermine the reliability and trustworthiness of machine learning systems by injecting hidden behaviors that can be maliciously ac...
By Juncheng Li, Yige Li, Hanxun Huang, Yunhao Chen, Xin Wang, Yixu Wang, Xingjun Ma, Yu-Gang Jiang
arXiv:2605.29655v4 Announce Type: replace
Abstract: Autoregressive multimodal large language models (MLLMs) enable 3D generation but struggle to scale to high-resolution shapes due to inadequate 3D t...
By Yuan Li, Congyi Zhang, Xifeng Gao, Xiaohu Guo
The paper presents a discrete generative model for neuronal spiking activity recorded on microelectrode arrays. It uses a shared vocabulary of spatiotemporal motifs learned by a residual vector‑quantized autoencoder and predicts motif occurrence with a factorized masked transformer. Evaluated on 31 assays from human brain organoids and ex vivo hippocampal tissue, the model achieves superior reconstruction and generation performance compared to baselines and shows that motifs are largely reused across assays.
By Md Sayed Tanveer, Mohammed A. Mostajo-Radji, Ge Wang
arXiv:2609.24057v1 Announce Type: cross
Abstract: Medical image interpretation is central to diagnosis and care, yet adapting general-purpose multimodal large language models (MLLMs) often requires r...
By Minda Zhao, Fangyu Hu, Yan Luo, Yutong Yang, Jiahui Cai, Kaichen Zhou, Manling Li, Paul Liang, Yilun Du, Lucy Q. Shen, Mengyu Wang
arXiv:2609.22241v1 Announce Type: new
Abstract: We present H2LooP Telecom Model v1, a domain-specialized large language models fine-tuned for the telecommunications industry. We release two domain-ad...
By Amit Singh, Vedant Nipane, Mayank Goel, Pulkit Agrawal, Sairanjan Mishra
arXiv:2609.22566v1 Announce Type: cross
Abstract: Knowledge distillation (KD) aims to compress high-performance teacher LLMs into lightweight students. However, distilled students often exhibit subst...
By Dileesha Kannangara, Sanghamitra Dutta
arXiv:2609.22603v1 Announce Type: new
Abstract: Summarization ships in countless production systems, making model selection a routine decision that depends on measuring summary quality. Existing metr...
By Nikhil Reddy Pottanigari, Ramin Fahimi, Noah Bolger, Sepideh Kharaghani, Ying Zhang
arXiv:2609.22793v1 Announce Type: new
Abstract: LLM-based machine translation evaluation can closely match human judgments, but in practice it remains largely diagnostic, with the signals rarely tran...
By Ji Hun Wang, Siyu Wu
arXiv:2609.22942v1 Announce Type: new
Abstract: Generative models are rapidly expanding image quality assessment (IQA) beyond traditional fidelity factors to emerging dimensions such as physical plau...
By Zhenchen Tang, Bo Peng, Zichuan Wang, Songlin Yang, Leilei Cao, Fengjie Zhu, Jing Dong
MolSC is a new dataset of 181,000 substituent-level examples that captures how attaching specific substituents to molecular scaffolds changes properties such as bioactivity and physicochemical descriptors. The authors also provide MolSC-Bench, a held‑out benchmark of 1,541 examples that are disjoint from MolSC at scaffold, substituent, and molecule levels. Experiments show that training molecular large language models on MolSC markedly improves their ability to predict substituent contributions, outperforming existing models on a range of downstream chemistry tasks.
By Hyuntae Park, Sooyeon Kim, Jiwon Park, SangKeun Lee
arXiv:2609.24691v1 Announce Type: new
Abstract: Latent diffusion models now dominate medical image generation, and every such pipeline rests on a \emph{tokenizer} that compresses images into the late...
By Niklas Bubeck, Yundi Zhang, Vasiliki Sideri-Lampretsa, Julian McGinnis, Jiancheng Yang, Daniel Rueckert, Jiazhen Pan
arXiv:2407.09693v3 Announce Type: replace
Abstract: The field of Neural-Symbolic (NeSy) systems is growing rapidly. Proposed approaches show great promise in achieving symbiotic unions of neural and...
By Charles Dickens, Connor Pryor, Changyu Gao, Alon Albalak, Eriq Augustine, William Wang, Stephen Wright, Lise Getoor