arXiv:2606. 14813v1 Announce Type: cross Abstract: Jet tagging at the Large Hadron Collider increasingly relies on deep learning models trained on massive simulated datasets, leading to high computational costs and limited robustness to detector mismodeling.
By Guillaume Letellier (LPCC), Antonin Vacheret (LPCC), Fr\'ed\'eric Jurie
arXiv:2601.11719v4 Announce Type: replace
Abstract: Self-supervised learning, in the context of foundation model training, is a powerful pre-training method for learning feature representations witho...
By Ho Fung Tsoi, Dylan Rankin
The paper introduces a data‑driven method for pairing events at the Large Hadron Collider using the energy mover's distance (EMD) to measure similarity, thereby creating augmentation‑free views for self‑supervised pre‑training. By matching distinct events based on EMD, the approach preserves the physics content of each event without handcrafted distortions. Experiments on QCD jets demonstrate that this pairing technique yields semantic jet embeddings with downstream discrimination power comparable to or better than traditional augmentation‑based baselines.
By Ho Fung Tsoi, Dylan Rankin
The paper compares convolutional neural networks (CNN), Vision Transformers (ViT), and hierarchical Swin Transformers for quark‑gluon jet classification using a three‑channel jet‑image representation. CNN and Swin models outperform ViT, indicating that local jet substructure is crucial for discrimination. The study also shows that block‑wise fine‑tuning, Momentum Contrast pretraining, and a compact Swin variant can improve performance while reducing parameters.
By Daeun Kim, Jaeyoon Cho, Jiwon Lee, Wonjun Jeong, Hyeongwoo Noh, Giyeong Kim, Seunghwan Yang, MinJung Kweon
The paper introduces a model calibration method using optimal transport to address discrepancies between simulation and experimental data in high-dimensional machine learning applications. Applied to jet tagging in particle physics, the technique calibrates a 128‑dimensional latent representation from a general‑purpose classifier, ensuring downstream derived quantities are properly calibrated. This enables more reliable use of foundation models for jet flavor analysis in LHC experiments and offers a general framework for correcting high‑dimensional simulations across scientific fields.
By Malte Algren, Tobias Golling, Francesco Armando Di Bello, Christopher Pollard
arXiv:2512. 07420v3 Announce Type: replace-cross Abstract: Jet identification plays a central role in analyzing data from high-energy collider experiments.
By Md Raqibul Islam, Adrita Khan, Mir Sazzat Hossain, Choudhury Ben Yamin Siddiqui, Md. Zakir Hossan, Tanjib Khan, M. Arshad Momen, Amin Ahsan Ali, AKM Mahbubur Rahman
arXiv:2609.06686v1 Announce Type: cross
Abstract: We present an unsupervised search for anomalous dijet events in proton--proton collision data using neural spline flow density estimation. A normaliz...
By Bhavishya Chebrolu (VIT-AP University, Amaravati, India), Hitesh Rasineni (VIT-AP University, Amaravati, India), Prajwal Aaryan Immadi (VIT-AP University, Amaravati, India)
arXiv:2606. 14373v1 Announce Type: cross Abstract: The workflow from particle collision to physics analysis passes through a series of reconstruction steps that are traditionally modular and disconnected, with no shared representation linking low-level detector data to high-level analysis tasks.
By Farouk Mokhtar, Joosep Pata, Michael Kagan, Javier Duarte
The paper introduces a particle‑level generative model that uses residual‑quantized full‑event data to enable fast, ML‑based surrogate simulation for collider events. It demonstrates conditional generation from detector‑stable particles, explores scaling across dataset and model sizes, and shows that token‑level loss predicts downstream physical fidelity. The work offers an empirical framework for scalable collider full‑event generation using residual‑quantized representations.
By Dan Godi, Dmitrii Kobylianskii, Eilam Gross
The paper introduces a Deep Sets surrogate for optimal transport (OT) that respects key metric properties—non-negativity, exchange symmetry, and zero self-distance—while leaving the triangle inequality unconstrained. Applied to the Energy Mover's Distance between collider events, the Metric-Aware Particle Flow Network achieves percent‑level mean absolute percentage error and markedly higher inference throughput compared to other exact and approximate methods. The architectural constraints also dramatically reduce triangle‑inequality violations, improving geometric fidelity across a large set of held‑out event triplets.
By Lauren Hay, Rishabh Jain, Matt LeBlanc, Jennifer Roloff
arXiv:2607. 24943v1 Announce Type: cross Abstract: In many classification problems, reliable instance-level labels are unavailable.
By Rapha\"el Bonnet-Guerrini, Johann Ioannou-Nikolaides, Troels Petersen, Vincenzo Piuri
arXiv:2607. 12726v1 Announce Type: cross Abstract: Global fits in high energy physics and cosmology often face the challenge of exploring high-dimensional parameter spaces with computationally expensive or topologically complex likelihood functions.
By Jorge Alda, Jacobo Asorey, Alejandro Mir, Siannah Pe\~naranda