The paper argues that language operates with two parameters: amplitude, which measures how often words co‑occur, and phase, a signed relational factor that determines how co‑activated meanings combine and can reverse a meaning’s contribution. Unlike amplitude, phase is not captured by standard word embeddings or transformer attention weights and is indexed to individuals and dyadic interactions. The authors propose six empirical predictions to test phase’s role and suggest that future language models should incorporate agent‑indexed, phase‑bearing semantic states.
arXiv:2606. 06238v1 Announce Type: new Abstract: We propose a statistical-field framework for text generated by large language models (LLMs), treating token embeddings as continuous spin variables on a one-dimensional chain.
By Huajian Ruan, Jinyang Li, Xingyu Guo, Lingxiao Wang
The paper introduces a framework called stochastic lexical calculus that determines when probabilities produced by large language models can be used to represent sequential states in scientific systems. It defines typed measurable transformations of contextual language, constructs a minimal closed representation, and provides necessary and sufficient conditions for unique semantic updates. The authors prove bounds on irreducible nonclosure and accumulated error, and show that under average contraction an external random recursion on a probability simplex is stable and unique. Empirical tests on frozen experiments demonstrate that raw prompt-conditioned probabilities fail an invariance gate, but after prompt-specific calibration a common three-state representation satisfies stability gates and covers 28 of 30 eight-step paths, achieving 0.933 coverage at a nominal 0.90 level.
By Matthew F Dixon
arXiv:2602.13194v3 Announce Type: replace-cross
Abstract: Humans and large language models can predict next letter or word from its prior context much better than random guessing, indicating strong r...
By Weishun Zhong, Doron Sivan, Tankut Can, Mikhail Katkov, Misha Tsodyks
arXiv:2606. 08417v1 Announce Type: cross Abstract: Diffusion and continuous flow-based language models have emerged as the leading non-autoregressive alternatives to language modeling.
By Antonio Franca, Alexander Tong
arXiv:2607. 07891v1 Announce Type: cross Abstract: Roy Harris's Integrationist linguistics offers a compelling critique of the referentialist tradition embedded deep at the heart of computational approaches to language, arguing that language is not a code that maps onto a pre-given world but a situated, bipartite activity oriented toward prospective joint action.
By J. Mark Bishop, Stephen J. Cowley