arXiv Machine Learning

Similarity Pairing with Energy Mover's Distance for Self-Supervised Pre-Training at the LHC

arXiv AI
Jun 16

JetParticle-JEPA: An Efficient Self-Supervised Representation Learning method for Jet Tagging in High-Energy Physics

arXiv:2606. 14813v1 Announce Type: cross Abstract: Jet tagging at the Large Hadron Collider increasingly relies on deep learning models trained on massive simulated datasets, leading to high computational costs and limited robustness to detector mismodeling.

By Guillaume Letellier (LPCC), Antonin Vacheret (LPCC), Fr\'ed\'eric Jurie
arXiv Machine Learning
4d ago

Learning the Geometry of Collider Events with Metric-Aware Deep Sets

The paper introduces a Deep Sets surrogate for optimal transport (OT) that respects key metric properties—non-negativity, exchange symmetry, and zero self-distance—while leaving the triangle inequality unconstrained. Applied to the Energy Mover's Distance between collider events, the Metric-Aware Particle Flow Network achieves percent‑level mean absolute percentage error and markedly higher inference throughput compared to other exact and approximate methods. The architectural constraints also dramatically reduce triangle‑inequality violations, improving geometric fidelity across a large set of held‑out event triplets.

By Lauren Hay, Rishabh Jain, Matt LeBlanc, Jennifer Roloff
arXiv Machine Learning
Jul 31

A Lightweight Foundation Model for Collider Physics with Multi-Domain Adaptation

arXiv:2607. 27501v1 Announce Type: new Abstract: We present a lightweight approach to foundation modeling (\textbf{NEXUS}) that leverages pre-trained learning from collider physics data towards out-of-domain tasks in other scientific datasets, using a fully connected autoencoder model with approximately 3 million parameters.

By Liangyu Wu, Qibin Liu, Alexander Yue, Julia Gonski
arXiv Machine Learning
Jun 8

ScatterPrism: convergence for generative simulation and inverse problems in particle and nuclear physics

arXiv:2604. 01313v2 Announce Type: replace Abstract: High-fidelity simulations and complex inverse problems, such as detector modeling and unfolding, are computationally intensive bottlenecks across subatomic physics, yet essential for accurate physical interpretation.

By Zeyu Xia, Tyler Kim, Trevor Reed, Judy Fox, Geoffrey Fox, Adam Szczepaniak
arXiv Machine Learning
Sep 10

Mind the Gap: Navigating Inference with Optimal Transport Maps

The paper introduces a model calibration method using optimal transport to address discrepancies between simulation and experimental data in high-dimensional machine learning applications. Applied to jet tagging in particle physics, the technique calibrates a 128‑dimensional latent representation from a general‑purpose classifier, ensuring downstream derived quantities are properly calibrated. This enables more reliable use of foundation models for jet flavor analysis in LHC experiments and offers a general framework for correcting high‑dimensional simulations across scientific fields.

By Malte Algren, Tobias Golling, Francesco Armando Di Bello, Christopher Pollard
arXiv Computer Vision
Aug 25

How Architecture and Training Affect TPC Representations Across Experiments

arXiv:2608.21756v1 Announce Type: cross Abstract: Deep-learning efforts have increasingly shifted toward foundation model approaches. In experimental physics, this allows models and learned represent...

By Tyler Wheeler, Michelle P. Kuchera, Raghuram Ramanujan, William Sieland, Ryan Krupp, Daniel Bazin, Connor L. Cross, Hoi Yan Ian Heung, Andrew J. Jones, Ruchi Mahajan, Saiprasad Ravishankar, Pranjal Singh, Benjamin Votaw, Chris Wrede
arXiv Computer Vision
Sep 2

Panda Diplomacy: Foundation Model Pre-training across Particle Imaging Detectors for High Energy and Nuclear Physics

Panda Diplomacy introduces a point‑cloud self‑distillation framework that enables a single foundation‑model architecture and objective to be pre‑trained across three distinct particle‑detector modalities—liquid argon time‑projection chambers, collider TPCs, and water Cherenkov detectors—without extensive modification. Using only 1,000 labeled images for downstream adaptation, the resulting Panda V2 model matches or surpasses specialized baselines that require orders of magnitude more supervision, achieving state‑of‑the‑art particle‑clustering performance with 70× fewer labeled events on sPHENIX and up to 1,000× fewer labels on LArTPC data. Linear probes further demonstrate that the model’s latent space captures physically meaningful structures such as particle causality and track curvature.

By Samuel Young, C\'esar Jes\'us-Valls, Kazuhiro Terao