Fine-tuning and adaptation

LoRA, PEFT, instruction tuning and domain adaptation — adapting a pretrained model without paying to train one.

3,783 stories · RSS feed

arXiv Machine Learning
Jun 30

On Surrogate Modeling of Static Response of AM Short-Fiber Thermoplastics Using Graph Neural Networks

arXiv:2606. 28996v1 Announce Type: new Abstract: Short-fiber thermoplastic (SFT) composites are increasingly employed in lightweight aerospace and automotive structures owing to their favorable strength-to-weight ratio, high production rates, and recyclability.

By Pharindra Pathak (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Vipin Kumar (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Trenton M. Ricks (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Suhasini Gururaja (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University), Siddhartha Srivastava (Auburn University, Oakridge National Lab, NASA Glenn Research Center, Auburn University, Auburn University)
arXiv AI
Jun 30

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

arXiv:2606. 30543v1 Announce Type: cross Abstract: With the proliferation of speech AI agents, understanding emotional entrainment in conversational interaction has become increasingly important.

By Sathvik Manikantan Napa Ugandhar, Hao Zhang, Alison Gunzler, Yuzhe Wang, Thomas Thebaud, Georgi Tinchev, Venkatesh Ravichandran, Laureano Moro-Vel\'azquez
arXiv AI
Jun 30

Animation2Code: Evaluating Temporal Visual Reasoning in Video-to-Code Generation

arXiv:2606. 28593v1 Announce Type: cross Abstract: While recent vision-language models (VLMs) have achieved significant improvements on static visual-to-code tasks such as generating code for webpages, charts, or SVGs, it remains unclear whether they can recover temporal dynamics when motion is present.

By Anya Ji, Abhijith Varma Mudunuri, David M. Chan, Alane Suhr
arXiv Machine Learning
Jun 30

HSAP: A Hierachical Sequence-aware Parallelism for Hybrid-Context Generative Models

arXiv:2606. 30460v1 Announce Type: new Abstract: In this paper, we aim to combine the advantages of existing sequence parallelism paradigms and overcomes their drawbacks, the most serious of which is the incapability to correctly compute causal attention on the hybrid-context packed sequences, in a stronger sequence parallelism framework.

By Songxin Zhang, Zejian Xie, Zhuoyang Song, Cong lin, Junyu Lu, Jiaxing Zhang, Bingyi Jing