arXiv Machine Learning

SPADE: Split-and-Delay Embeddings for Autoregressive High-Granularity Calorimeter Simulation

arXiv:2606. 11304v1 Announce Type: cross Abstract: We introduce SPADE (SPlit And Delay Embeddings), an autoregressive transformer for sequences whose tokens carry multiple features.

arXiv Machine Learning
Jun 16

GPT-Based Fast Simulation of CLAS12 Detector Hits via Conditional Autoregressive Generation

arXiv:2606. 16035v1 Announce Type: cross Abstract: Modern particles physics experiments have demonstrated an increasing need for fast, high-fidelity detector simulation as detector components have improved and subsequent computational requirements approach the limits of available resources.

By Cole Granger, James Giroux, Richard Tyson, Maurizio Ungaro, Cristiano Fanelli
arXiv Machine Learning
Jun 4

CaloTrilogy: Toward a Breakthrough in One-Step, End-to-End, Physics-Guided Shower Generation for Modern Calorimeters

arXiv:2606. 04165v1 Announce Type: cross Abstract: High-precision calorimeter simulation at current and future colliders imposes rapidly growing computational demands, motivating the development of machine-learning surrogates for traditional Monte Carlo tools such as Geant4.

By Cheng Jiang, Sitian Qian, Kevin Pedro, Oz Amram, Huilin Qu, Maggie Voetberg
arXiv Machine Learning
1d ago

Scaling Collider Event Generation with Residual-Quantized Tokens

The paper introduces a particle‑level generative model that uses residual‑quantized full‑event data to enable fast, ML‑based surrogate simulation for collider events. It demonstrates conditional generation from detector‑stable particles, explores scaling across dataset and model sizes, and shows that token‑level loss predicts downstream physical fidelity. The work offers an empirical framework for scalable collider full‑event generation using residual‑quantized representations.

By Dan Godi, Dmitrii Kobylianskii, Eilam Gross
arXiv AI
Jun 16

JetParticle-JEPA: An Efficient Self-Supervised Representation Learning method for Jet Tagging in High-Energy Physics

arXiv:2606. 14813v1 Announce Type: cross Abstract: Jet tagging at the Large Hadron Collider increasingly relies on deep learning models trained on massive simulated datasets, leading to high computational costs and limited robustness to detector mismodeling.

By Guillaume Letellier (LPCC), Antonin Vacheret (LPCC), Fr\'ed\'eric Jurie
arXiv Computer Vision
Sep 2

Panda Diplomacy: Foundation Model Pre-training across Particle Imaging Detectors for High Energy and Nuclear Physics

Panda Diplomacy introduces a point‑cloud self‑distillation framework that enables a single foundation‑model architecture and objective to be pre‑trained across three distinct particle‑detector modalities—liquid argon time‑projection chambers, collider TPCs, and water Cherenkov detectors—without extensive modification. Using only 1,000 labeled images for downstream adaptation, the resulting Panda V2 model matches or surpasses specialized baselines that require orders of magnitude more supervision, achieving state‑of‑the‑art particle‑clustering performance with 70× fewer labeled events on sPHENIX and up to 1,000× fewer labels on LArTPC data. Linear probes further demonstrate that the model’s latent space captures physically meaningful structures such as particle causality and track curvature.

By Samuel Young, C\'esar Jes\'us-Valls, Kazuhiro Terao
arXiv Machine Learning
Sep 17

Similarity Pairing with Energy Mover's Distance for Self-Supervised Pre-Training at the LHC

The paper introduces a data‑driven method for pairing events at the Large Hadron Collider using the energy mover's distance (EMD) to measure similarity, thereby creating augmentation‑free views for self‑supervised pre‑training. By matching distinct events based on EMD, the approach preserves the physics content of each event without handcrafted distortions. Experiments on QCD jets demonstrate that this pairing technique yields semantic jet embeddings with downstream discrimination power comparable to or better than traditional augmentation‑based baselines.

By Ho Fung Tsoi, Dylan Rankin