arXiv AI

CAS I: A Geometric Coding Theorem

arXiv:2607. 13796v1 Announce Type: cross Abstract: This paper establishes a direct analogue of the classical Coding Theorem in the setting of symmetry groups.

arXiv AI
Jul 14

Instruction Set and Language for Hypergraphs

arXiv:2607. 10194v1 Announce Type: cross Abstract: We present IsalHG, a method for representing the structure of any finite, connected hypergraph of bounded hyperedge arity as a string over a compact instruction alphabet $\Sigma_{\mathrm{HG}}$.

By Mario Pascual-Gonzalez, Ezequiel Lopez-Rubio
arXiv Machine Learning
1d ago

Group-Invariant Statistics Determine Embedding Geometry: Harmonic Analysis of Representations from Bach to the Night Sky

The paper shows that the geometric patterns seen in language model embeddings—such as circles for months and saddle-shaped manifolds—arise from group-invariant statistics in word co‑occurrence data. By extending previous work on translation symmetry to arbitrary finite, compact, and homogeneous groups, the authors prove that embeddings correspond to matrix elements of the irreducible representations of the symmetry group. They validate this theory experimentally with the cyclic group <Z_{12}> for months, a dihedral group for musical chords, and a spherical‑harmonic embedding for celestial objects.

By Liam Storan, Andreas Tolias, Nina Miolane
arXiv AI
Jul 24

Representative Sets in Propositional Abduction

arXiv:2607. 21183v1 Announce Type: cross Abstract: The propositional abduction problem is a well-known form of non-monotonic reasoning where we are asked to find an explanation of a given manifestation.

By Johannes Schmidt (J\"onk\"oping University), Mohamed Maizia (J\"onk\"oping University, Link\"oping University), Victor Lagerkvist (Link\"oping University), Johannes K. Fichte (Link\"oping University)
arXiv Machine Learning
Jun 2

Symmetries in PAC-Bayesian Learning

arXiv:2510. 17303v2 Announce Type: replace Abstract: Symmetries are known to improve the empirical performance of machine learning models, yet theoretical guarantees explaining these gains remain limited.

By Armin Beck, Peter Ochs
arXiv Machine Learning
Jul 10

High-Dimensional Procrustes Matching via Tree Counts

arXiv:2607. 08538v1 Announce Type: cross Abstract: Suppose we observe two sets of $n$ Gaussian vectors in $\mathbb{R}^d$, with the promise that, after applying a permutation of $[n]$ and a rotation of $\mathbb{R}^d$, the two sets are $\rho$-correlated.

By Xiaochun Niu, Tselil Schramm, Jiaming Xu
arXiv AI
Aug 14

Algebraic Decomposition Theory for Transformer Length Generalization

arXiv:2608. 13433v1 Announce Type: cross Abstract: Transformer-based language models are known to sometimes generalize to sequences longer than seen during training, but we lack a precise characterization of which tasks admit length generalization.

By Andy Yang, Blerta Veseli, Corentin Barloy, Micha\"el Cadilhac, Andreas Krebs, Charles Paperman, Howard Straubing, Michael Hahn