arXiv:2606. 14870v1 Announce Type: cross Abstract: Foundation models (FMs) trained on large datasets and fine-tuned on downstream tasks have emerged as a powerful paradigm in AI for science.
By Ibrahim Elsharkawy, Joschka Birk, Vinicius Mikuni, Wahid Bhimji, Gregor Kasieczka, Benjamin Nachman
arXiv:2606. 14813v1 Announce Type: cross Abstract: Jet tagging at the Large Hadron Collider increasingly relies on deep learning models trained on massive simulated datasets, leading to high computational costs and limited robustness to detector mismodeling.
By Guillaume Letellier (LPCC), Antonin Vacheret (LPCC), Fr\'ed\'eric Jurie
arXiv:2606. 19781v1 Announce Type: cross Abstract: Neural scaling laws describe how model performance improves as a power law in compute, model size, and dataset size.
By Jan-Lucas Uslu, Kevin Greif, Daniel Whiteson, Benjamin Nachman
The paper investigates whether two particle physics foundation models, OmniLearned and ParticleViT, pretrained on high‑energy proton–proton and electron–proton collisions can transfer knowledge to a low‑energy neutrino experiment. Using MINERvA neutrino–nucleus scattering data, the authors evaluate the models on energy regression and charged‑current pion classification tasks, finding that the pretrained models outperform similarly sized models trained from scratch, with OmniLearned excelling in regression and ParticleViT in classification. When the same transformer architecture is initialized from unrelated text pretraining (BERT), the performance advantage is minimal for classification and absent for regression, indicating that particle‑level foundation models capture inductive biases that generalize across energy scales, detector technologies, and physics processes.
By Gregor Krzmanc, Vinicius Mikuni, Benjamin Nachman, Callum Wilkinson
arXiv:2606. 14373v1 Announce Type: cross Abstract: The workflow from particle collision to physics analysis passes through a series of reconstruction steps that are traditionally modular and disconnected, with no shared representation linking low-level detector data to high-level analysis tasks.
By Farouk Mokhtar, Joosep Pata, Michael Kagan, Javier Duarte
arXiv:2607. 23377v1 Announce Type: cross Abstract: The largest machine learning models in particle physics are also the most expensive to train, yet the return on scaling a given architecture cannot be estimated before that compute is spent.
By Jan-Lucas Uslu, Benjamin Nachman, Christopher Re