Fine-tuning and adaptation

LoRA, PEFT, instruction tuning and domain adaptation — adapting a pretrained model without paying to train one.

3,721 stories · RSS feed

arXiv AI
Jun 30

McMg: A Learned Phase-Space Multi-channel Multigrid Preconditioner for Helmholtz Equation

arXiv:2606. 30495v1 Announce Type: cross Abstract: Solving heterogeneous Helmholtz equations at high wavenumbers remains challenging because the discretized operator is indefinite, pollution degrades phase accuracy, and scalar coarse-grid correction can discard the local phase and propagation-direction information carried by oscillatory errors.

By Jiwei Jia, Xinliang Liu, Juntao Wang, Jinchao Xu
arXiv Machine Learning
Jun 30

On Surrogate Modeling of Static Response of AM Short-Fiber Thermoplastics Using Graph Neural Networks

arXiv:2606. 28996v1 Announce Type: new Abstract: Short-fiber thermoplastic (SFT) composites are increasingly employed in lightweight aerospace and automotive structures owing to their favorable strength-to-weight ratio, high production rates, and recyclability.

By Pharindra Pathak (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Vipin Kumar (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Trenton M. Ricks (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Suhasini Gururaja (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Siddhartha Srivastava (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University)
arXiv AI
Jun 30

Dockerless: Environment-Free Program Verifier for Coding Agents

arXiv:2606. 28436v1 Announce Type: cross Abstract: Program verifiers play a central role in training coding agents, including selecting trajectories for supervised fine-tuning (SFT) and providing rewards for reinforcement learning (RL).

By Wenhao Zeng, Yuling Shi, Xiaodong Gu, Chao Hu, Chaofan Wang, Yuhao Cui, Hongting Zhou, Mengnan Qi, Jianqiao Wangni, Zhaojian Yu, Shuzheng Gao, Kai Cai, Shilin He
arXiv AI
Jun 30

Animation2Code: Evaluating Temporal Visual Reasoning in Video-to-Code Generation

arXiv:2606. 28593v1 Announce Type: cross Abstract: While recent vision-language models (VLMs) have achieved significant improvements on static visual-to-code tasks such as generating code for webpages, charts, or SVGs, it remains unclear whether they can recover temporal dynamics when motion is present.

By Anya Ji, Abhijith Varma Mudunuri, David M. Chan, Alane Suhr
arXiv Machine Learning
Jun 30

HSAP: A Hierachical Sequence-aware Parallelism for Hybrid-Context Generative Models

arXiv:2606. 30460v1 Announce Type: new Abstract: In this paper, we aim to combine the advantages of existing sequence parallelism paradigms and overcomes their drawbacks, the most serious of which is the incapability to correctly compute causal attention on the hybrid-context packed sequences, in a stronger sequence parallelism framework.

By Songxin Zhang, Zejian Xie, Zhuoyang Song, Cong lin, Junyu Lu, Jiaxing Zhang, Bingyi Jing