arXiv Machine Learning By Jian Xu, Chao Li, Delu Zeng, John Paisley, Qibin Zhao

Higher-Order Token Interactions via Quantum Attention

Read the original on arXiv Machine Learning →

arXiv:2606. 11673v1 Announce Type: cross Abstract: Standard dot-product self-attention computes, in a single layer, only pairwise (order-2) interactions between tokens; representing a generic order-$k$ interaction is known to require either super-quadratic resources in one layer or composition across depth.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Jun 24

Quantum Adaptive Self-Attention for Quantum Transformer Models

arXiv:2504. 05336v4 Announce Type: replace-cross Abstract: A recurring weakness in quantum machine learning (QML) is that reported ``quantum advantages'' are seldom tested against a \emph{capacity-matched} classical control, leaving it unclear whether a gain comes from the quantum substrate or from the architectural change that accompanies it.

By Chi-Sheng Chen, En-Jui Kuo
arXiv Machine Learning
Sep 22

Watching Quantum Models Think: Hilbert-Space Interpretability in Quantum Transformer Blocks

The paper demonstrates that quantum transformer blocks can be intrinsically interpretable by tracking quantum mutual information, entanglement entropy, and state fidelity across layers. Experiments on four synthetic tasks show that learned mutual information aligns with task structure, entanglement is essential for accuracy, and mutual information predicts prediction correctness. These findings are validated on IBM Quantum hardware, illustrating that quantum computation’s physics can provide observable interpretability signals.

By Diego Iacopetta, Andrea Gasparini