arXiv Machine Learning
Jun 30

Atompack: A Storage and Distribution Layer for Read-Heavy Atomistic ML Training Datasets

arXiv:2606. 29975v1 Announce Type: new Abstract: Atomistic machine learning datasets are increasingly used for training: large immutable snapshots are read repeatedly, shuffled across epochs, staged across clusters' storage systems, and republished as reusable scientific artifacts.

By Ali Ramlaoui, Daniel T. Speckhard, Sagar Pal, Fragkiskos D. Malliaros, Alexandre Duval, Victor Schmidt
arXiv Machine Learning
Sep 14

Fast BIB simulation at a future Muon Collider with generative machine learning

The paper presents the first machine learning models for fast generation of beam‑induced background (BIB) in tracking detectors at a future Muon Collider. Two architectures are explored: a high‑fidelity tabular diffusion model and a faster circular spline flow model. Both produce BIB hits and tracks that closely match full simulation results, achieving over an order of magnitude speed‑up while requiring far less computational resources.

By Radha Mastandrea, Shiyu Peng, Benjamin Rosser, Matt LeBlanc