arXiv Machine Learning By Andrew Fitzgibbon, Christoph M. Wintersteiger, Jeffrey Sarnoff

Novel Aspects of IEEE SA P3109 Arithmetic Formats for Machine Learning

Read the original on arXiv Machine Learning →

arXiv:2606. 04028v1 Announce Type: new Abstract: The IEEE P3109 draft standard defines a parameterized family of binary floating-point formats and associated operations, with a focus on facilitating machine learning.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 7

Golden Ruler: A Numeric Format Catalog with Bit-Exact Conformance Vectors for FP8, BF16, MXFP4, and Microscaling Formats

The paper introduces the Golden Ruler, a catalog of 109 numeric formats for machine learning hardware, including FP8, BF16, MXFP4, and microscaling block formats. It provides six bit‑exact conformance packs that map to IEEE P3109 v3.2.0 standards, each as a self‑contained JSON document with a SHA‑256 fingerprint and an anchor vector for cross‑pack sanity checks. The work cross‑validates these packs against ml_dtypes 0.5.4 and documents any divergences as spec‑permitted gaps, offering a vendor‑neutral reference for engineers.

By Dmitrii Vasilev
arXiv Machine Learning
Jun 2

Stochastic Rounding Increases Small Singular Values

arXiv:2606. 00312v1 Announce Type: cross Abstract: Over the past half-dozen years, stochastic rounding (SR) has regained significant attention as a quantization scheme for low-precision floating-point arithmetic, with applications spanning numerical analysis and modern machine learning systems.

By Linkai Ma, Tingzhou Yu, Petros Drineas
arXiv AI
Jun 4

dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats

arXiv:2606. 04115v1 Announce Type: cross Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-width uniformly across all layers is sub-optimal in terms of both performance and accuracy.

By Giuseppe Franco, Ian Colbert, Pablo Monteagudo-Lago, Felix Marty, Nicholas Fraser