Large language models

Model releases, architecture work and prompting research on large language models — from frontier-lab announcements to the arXiv papers behind them.

18,142 stories · RSS feed

arXiv Machine Learning
Aug 7

A Six-Dimensional Taxonomy of Post-Training Adaptation Techniques with Applications in AI Governance

arXiv:2608. 06246v1 Announce Type: new Abstract: Post-training adaptation has become central to modern machine learning practice and includes techniques such as retraining, fine-tuning, parameter-efficient adaptation, alignment, retrieval augmentation, model editing, unlearning, calibration, and Multimodal Instruction Tuning.

By Fardin Afdideh, Fernando Seoane, Farhad Abtahi
arXiv AI
Aug 7

DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data

arXiv:2608. 05375v1 Announce Type: new Abstract: Clinical machine learning (ML) has the potential to support high-stakes medical decision-making, but reliable deployment is often constrained by scarce, heterogeneous, and temporal complexity.

By Ruilin Wang, Bo-Hong Wang, Elizabeth Kourbatski, Jun Bai, Hegang Chen, Ziyang Song, Gilles Boire, Marie Hudson, Yue Li