arXiv Machine Learning By Siqi Miao, Shitij Govil, Jack P. Rodgers, Mia Liu, Javier Duarte, Shih-Chieh Hsu, Yuan-Tang Chou, Pan Li

HEPTv2: End-to-End Efficient Point Transformer for Charged Particle Reconstruction

Read the original on arXiv Machine Learning →

arXiv:2606. 20437v1 Announce Type: cross Abstract: Charged-particle tracking -- reconstructing trajectories from sparse detector measurements -- is a fundamental high-energy-physics inference problem and a canonical example of learning under extreme combinatorial ambiguity.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 14

Learning the Geometry of Collider Events with Metric-Aware Deep Sets

The paper introduces a Deep Sets surrogate for optimal transport (OT) that respects key metric properties—non-negativity, exchange symmetry, and zero self-distance—while leaving the triangle inequality unconstrained. Applied to the Energy Mover's Distance between collider events, the Metric-Aware Particle Flow Network achieves percent‑level mean absolute percentage error and markedly higher inference throughput compared to other exact and approximate methods. The architectural constraints also dramatically reduce triangle‑inequality violations, improving geometric fidelity across a large set of held‑out event triplets.

By Lauren Hay, Rishabh Jain, Matt LeBlanc, Jennifer Roloff
arXiv Machine Learning
Sep 17

Comprehensive reconstruction of collider events with hypergraph representation learning and graph-conditioned diffusion

VyPER is a new geometric learning framework that reconstructs particle collider events by representing them as hypergraphs with a physics-inspired topology. It tackles two key tasks: assigning measured jets and leptons to their parent particles through supervised hyperedge classification, and predicting neutrino kinematics using a diffusion model, all optimized jointly with a shared loss function. The authors evaluate VyPER on various proton‑proton collision processes, showing improved performance over existing analytical and machine‑learning methods and enabling more precise measurements in Higgs, electroweak, and top‑quark studies.

By Lining Mao, Yvonne Peters, Ethan Simpson, Zihan Zhang
arXiv AI
Jun 16

JetParticle-JEPA: An Efficient Self-Supervised Representation Learning method for Jet Tagging in High-Energy Physics

arXiv:2606. 14813v1 Announce Type: cross Abstract: Jet tagging at the Large Hadron Collider increasingly relies on deep learning models trained on massive simulated datasets, leading to high computational costs and limited robustness to detector mismodeling.

By Guillaume Letellier (LPCC), Antonin Vacheret (LPCC), Fr\'ed\'eric Jurie
arXiv Machine Learning
Sep 25

Reconstructing short-lived particles using hypergraph representation learning

The paper introduces HyPER, a hypergraph-based graph neural network architecture designed to reconstruct short-lived particles in collider experiments. By leveraging hypergraph representation learning, HyPER builds more powerful and efficient representations of collider events, enabling accurate reconstruction of parent particles from final-state objects. In simulations, HyPER outperforms existing state‑of‑the‑art techniques while using fewer parameters, and its flexible hypergraph approach can be applied to a wide range of physics processes.

By Callum Birch-Sykes, Brian Le, Yvonne Peters, Ethan Simpson, Zihan Zhang
arXiv Computer Vision
Sep 2

Panda Diplomacy: Foundation Model Pre-training across Particle Imaging Detectors for High Energy and Nuclear Physics

Panda Diplomacy introduces a point‑cloud self‑distillation framework that enables a single foundation‑model architecture and objective to be pre‑trained across three distinct particle‑detector modalities—liquid argon time‑projection chambers, collider TPCs, and water Cherenkov detectors—without extensive modification. Using only 1,000 labeled images for downstream adaptation, the resulting Panda V2 model matches or surpasses specialized baselines that require orders of magnitude more supervision, achieving state‑of‑the‑art particle‑clustering performance with 70× fewer labeled events on sPHENIX and up to 1,000× fewer labels on LArTPC data. Linear probes further demonstrate that the model’s latent space captures physically meaningful structures such as particle causality and track curvature.

By Samuel Young, C\'esar Jes\'us-Valls, Kazuhiro Terao