Hugging Face Trending Papers

Scaling Up Thermodynamic AI Models

Read the original on Hugging Face Trending Papers →

Thermodynamic computing devices based on the Ising model show great promise for low-power AI inference and edge computing, but scalable methods for training large models for such hardware remain limited. Prior theory shows that the time-averaged behavior of high-temperature Gibbs-sampled Ising systems can implement feed-forward neural inference.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.

arXiv AI
Jul 2

Scaling Up Thermodynamic AI Models

arXiv:2607. 00170v1 Announce Type: cross Abstract: Thermodynamic computing devices based on the Ising model show great promise for low-power AI inference and edge computing, but scalable methods for training large models for such hardware remain limited.

By Andrew G. Moore
arXiv AI
Jun 9

Optimizing Energy-based Neural Network Training with Coherent Ising Machine

arXiv:2606. 09117v1 Announce Type: cross Abstract: While Ising machines serve as advanced physical solvers for the Ising model,enabling applications in combinatorial optimization and neural network training,their scalability for large-scale neural networks remains constrained by hardware connectivity limitations and suboptimal training methodologies.

By Chen-Rui Fan, Bo Lu, Zhi-Hong Zhang, Run-Qing Zhang, Jing-Wei Wen, Chuan Wang
Hugging Face Trending Papers
Jul 29

Equilibrium Training of Energy-Based Models with Parallel Trajectory Tempering

Energy-Based Models (EBMs) provide an interpretable framework for generative modeling of scientific data, but poor Markov Chain Monte Carlo mixing often limits their reliability. We introduce a training algorithm based on Parallel Trajectory Tempering (PTT), which exploits the continuity of the optimization path to maintain equilibrium sampling throughout learning.

arXiv Machine Learning
Aug 27

Thermodynamic cost of inference and learning in physical neural networks

The paper investigates the thermodynamic cost of inference and learning in physical neural networks. It shows that quasi‑static inference requires no work, while finite‑speed inference incurs work bounded by the Wasserstein‑2 distance between thermal states, roughly $k_B T$ per dimension of the widest layer. Learning, however, has an irreducible cost of a few $k_B T$ per parameter, independent of speed, indicating that memory dominates the thermodynamic price.

By Alexei V. Tkachenko
arXiv Machine Learning
Jul 20

A Blueprint for Equilibrium-Based Differentiable Continuous-Variable Thermodynamic Computing

arXiv:2607. 16183v1 Announce Type: new Abstract: To address the escalating energy and latency demands of machine-learning workloads, we introduce a blueprint for an energy-efficient and fast thermodynamic computing stack that leverages stochastic analog processes in physical hardware.

By Owen Lockwood, J\'er\'emy B\'ejanin, Joost Bus, Christopher Chamberland, Patrick Huembeli, Frank Sch\"afer, Guillaume Verdon