We develop a convergent scheme to train neural networks involving analytic activation functions based on gradient flows. Convergence properties are guaranteed by Lojasiewicz theory.
arXiv:2601. 07397v2 Announce Type: replace-cross Abstract: In this work, we propose a novel layerwise adaptive construction method for neural network architectures.
By Michael Hinterm\"uller, Michael Hinze, Denis Korolev
arXiv:2607. 24726v1 Announce Type: new Abstract: The Deep Galerkin Method (DGM) and Physics Informed Neural Networks (PINNs) have become widely-used methods for solving partial differential equations (PDEs) in the rapidly growing field of scientific machine learning.
By Justin Sirignano, Konstantinos Spiliopoulos, Samuel Cohen
arXiv:2402. 00152v5 Announce Type: replace Abstract: Constructing the architecture of a neural network is a challenging pursuit for the machine learning community, and the dilemma of whether to go deeper or wider remains a persistent question.
By Yahong Yang, Juncai He
arXiv:2507. 14177v2 Announce Type: replace-cross Abstract: This paper aims to understand the training solution, which is obtained by the back-propagation algorithm, of two-layer neural networks whose hidden layer is composed of the units with smooth activation functions, including the usual sigmoid type most commonly used before the advent of ReLUs.
By Changcun Huang
arXiv:2505. 12430v2 Announce Type: replace Abstract: Recently, innovative adaptations of the Ritz method incorporating deep learning have been developed, known as the Deep Ritz Method.
By Rafael Florencio, Julio Guerrero
arXiv:2606. 12337v1 Announce Type: cross Abstract: Inverse problems governed by partial differential equations (PDEs) are central to computational mechanics and are commonly solved by adjoint-based optimization, while physics-informed neural networks (PINNs) have emerged as a flexible alternative.
By Zhen Zhang, Alessandro Alla, George Em Karniadakis
arXiv:2607. 10200v1 Announce Type: new Abstract: The Neural Tangent Kernel (NTK) is one powerful tool for analyzing the training dynamics of neural networks in the over-parameterized regime.
By Bangti Jin, Longjun Wu
arXiv:2608. 14733v1 Announce Type: cross Abstract: Building on the foundation of single-hidden-layer neural networks, Fourier Feature Networks (FENs) are proposed, which incorporate Fourier features using $\cos$, $\sin$, or a combination of both.
By Qihong Yang, Zhijie Su, Yangtao Deng, Qiaolin He
arXiv:2608. 16475v1 Announce Type: cross Abstract: The Porous Medium Equation (PME), given by $u_t = \Delta(u^m)$ for $m > 1$, is a degenerate nonlinear parabolic partial differential equation that arises in various physical applications such as fluid flow in porous media, heat transfer in plasmas, and population dynamics.
By Noura Al Helwani, Sophie Moufawad, Nabil Nassif
arXiv:2607. 02194v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a promising route to solve partial differential equations, yet they have struggled to reach the precision of classical solvers.
By Joseph Webb, Sadok Jerad, Coralia Cartis
arXiv:2607. 02003v1 Announce Type: cross Abstract: Although neural networks are remarkably effective, their underlying optimization principles remain theoretically elusive, often characterized by non-convex landscapes and stochastic heuristics.
By Matej Benko, Pierre Bousquet, Iwona Chlebicka, B{\l}a\.zej Miasojedow