Hugging Face Trending Papers

A Study of Hidden-State Optimization Order in Predictive Coding Networks

arXiv AI
Sep 2

A Study of Hidden-State Optimization Order in Predictive Coding Networks

The paper investigates how the sequence of hidden-state optimization affects feature learning in local-learning models, specifically predictive coding networks (PCNs). It introduces a boundary-first inference schedule that first aligns hidden states at chunk boundaries before refining representations within each chunk. Experiments on CIFAR-10 show that this approach improves accuracy by 9.77% over standard PCNs and 5.51% under a different parametrization, with diagnostics indicating stronger feature learning.

By Xueyuan Li, Danilo Vasconcellos Vargas
Hugging Face Trending Papers
Aug 3

Topological Simplification in Predictive Coding Networks

We study the topology of learned representations in predictive coding networks (PCNs), a neuro-inspired bidirectional architecture, using a quantitative layer-wise persistent homology analysis. We train well-performing PCNs on a synthetic classification dataset ($\geq 99.

Hugging Face Trending Papers
Jul 2

DRDN: Decoupled Representation Dynamic Network for From-Scratch ViT Class-Incremental Learning

Dynamic expansion methods for class-incremental learning (CIL) protect task-specific knowledge by growing dedicated tokens or subnetworks, yet our analyses suggest that classification supervision alone does not sufficiently preserve task-agnostic shared backbone representations over long incremental sequences. We identify two intertwined challenges: cross-task confusion from sequential training on predominantly current-task data, which biases decision boundaries toward recent tasks; and under-optimized shared representations in the backbone that cap long-term discriminability as tasks accumulate.

arXiv AI
Sep 21

Neural Cellular Automata Learn General Features in their Hidden Channels

Neural Cellular Automata (NCAs) are shown to learn general, scale‑invariant topological primitives in their hidden channels, which can be transferred from a teacher to a student model for few‑shot learning. The study introduces a transfer‑learning mechanism that injects pretrained hidden states into a student, improving early optimization and outperforming recurrent and feed‑forward baselines on MNIST benchmarks with only ~9,800 parameters. Mechanistic analysis reveals that hidden channels decouple feature extraction from classification, converging to mutually orthogonal states that absorb morphological complexity.

By Etienne Guichard, Stefano Nichele