Hugging Face Blog

Simple considerations for simple people building fancy neural networks

arXiv AI
Jul 9

Understanding Two-Layer Neural Networks with Smooth Activation Functions

arXiv:2507. 14177v2 Announce Type: replace-cross Abstract: This paper aims to understand the training solution, which is obtained by the back-propagation algorithm, of two-layer neural networks whose hidden layer is composed of the units with smooth activation functions, including the usual sigmoid type most commonly used before the advent of ReLUs.

By Changcun Huang
arXiv Machine Learning
Sep 11

Revenge of Monosemanticity: Neuron Specialization as a New Form of Feature Learning in MLPs

The paper investigates how multilayer perceptrons (MLPs) learn features in regression tasks with clustered data. It finds that instead of forming a single global low‑dimensional representation, MLPs develop monosemantic specialized neurons—each neuron aligns strongly with a specific predictive feature relevant to a particular region of the input space. This specialization results in a collection of local low‑dimensional representations, giving MLPs a provable data‑efficiency advantage over methods that rely on a global representation.

By Amirhesam Abedsoltan, Enric Boix-Adsera, Fivos Kalogiannis, Mikhail Belkin
Hugging Face Trending Papers
Jun 24

Variational Autoencoder Layer

Variational Autoencoders (VAEs) belong to a family of autoencoders with probabilistic properties, making them well suited for generating data by producing a smooth and continuous latent space. Despite being introduced over a decade ago, the method continues to be widely adopted in both research and industry for diverse applications.