arXiv Machine Learning

Higher-Order Token Interactions via Quantum Attention

arXiv:2606. 11673v1 Announce Type: cross Abstract: Standard dot-product self-attention computes, in a single layer, only pairwise (order-2) interactions between tokens; representing a generic order-$k$ interaction is known to require either super-quadratic resources in one layer or composition across depth.

arXiv Machine Learning
Jun 24

Quantum Adaptive Self-Attention for Quantum Transformer Models

arXiv:2504. 05336v4 Announce Type: replace-cross Abstract: A recurring weakness in quantum machine learning (QML) is that reported ``quantum advantages'' are seldom tested against a \emph{capacity-matched} classical control, leaving it unclear whether a gain comes from the quantum substrate or from the architectural change that accompanies it.

By Chi-Sheng Chen, En-Jui Kuo
arXiv Machine Learning
Sep 22

Watching Quantum Models Think: Hilbert-Space Interpretability in Quantum Transformer Blocks

The paper demonstrates that quantum transformer blocks can be intrinsically interpretable by tracking quantum mutual information, entanglement entropy, and state fidelity across layers. Experiments on four synthetic tasks show that learned mutual information aligns with task structure, entanglement is essential for accuracy, and mutual information predicts prediction correctness. These findings are validated on IBM Quantum hardware, illustrating that quantum computation’s physics can provide observable interpretability signals.

By Diego Iacopetta, Andrea Gasparini
arXiv Machine Learning
Aug 11

Classical $\mathrm{SU}(2)$ Models Match or Exceed Shallow Variational Quantum Circuits on Vision Benchmarks

arXiv:2608. 07822v1 Announce Type: cross Abstract: Quaternion-valued neural networks and variational quantum circuits (VQCs) both derive local transformations from $\mathrm{SU}(2)$ geometry, yet their performance on classical supervised learning remains poorly understood.

By Christopher Fulton, Irene Tsapara, Lawrence Fulton