arXiv Machine Learning

Singular perturbations and hierarchical learning in two-layer neural networks

arXiv:2607. 10869v1 Announce Type: new Abstract: We study the population gradient flow of an infinitely wide two-layer neural network learning a misspecified single-index model in high dimension.

arXiv AI
Jul 9

Understanding Two-Layer Neural Networks with Smooth Activation Functions

arXiv:2507. 14177v2 Announce Type: replace-cross Abstract: This paper aims to understand the training solution, which is obtained by the back-propagation algorithm, of two-layer neural networks whose hidden layer is composed of the units with smooth activation functions, including the usual sigmoid type most commonly used before the advent of ReLUs.

By Changcun Huang