arXiv:2607. 27000v1 Announce Type: cross Abstract: Optimization in non-convex neural network models is strongly influenced by the geometry of the solution space: sparse, isolated, point-like clusters are typically algorithmically inaccessible, whereas wide and flat regions can be found efficiently despite being relatively rare.
By Enrico M. Malatesta, Alessandra Passalacqua, Riccardo Zecchina
arXiv:2608. 11911v1 Announce Type: cross Abstract: A central promise of useful quantum advantage is the ability to compute ground states of Hamiltonian systems beyond the reach of classical simulation methods.
By Timothy Heightman, Elena Orlova, Philip Mantrov, Aleksei Ustimenko
arXiv:2607. 11555v1 Announce Type: new Abstract: Learning neural set functions is pivotal to a wide range of important applications, including compound selection in AI-driven drug discovery and product recommendation.
By Yongquan Shi, Zijing Ou, Shiping Wang, Yatao Bian
arXiv:2606. 22984v2 Announce Type: replace-cross Abstract: Efficient sampling of the Boltzmann distribution in frustrated spin glasses is central to statistical mechanics and combinatorial optimization.
By Lu Zhong, Wenli Duan, Jing Liu, Pan Zhang, Ying Tang
arXiv:2412. 04177v2 Announce Type: replace Abstract: Recently, there has been an increasing interest in performing post-hoc uncertainty estimation about the predictions of pre-trained deep neural networks (DNNs).
By Luis A. Ortega, Sim\'on Rodr\'iguez-Santana, Daniel Hern\'andez-Lobato
arXiv:2608. 16760v1 Announce Type: new Abstract: Reliable optimization is central to neural network (NN) training, yet Adam, the default optimizer for modern LLMs, rests on a fragile foundation.
By Yushun Zhang
arXiv:2206. 04359v3 Announce Type: replace Abstract: One of the fundamental challenges in the deep learning community is to theoretically understand how well a deep neural network generalizes to unseen data.
By Chengli Tan, Jiangshe Zhang, Junmin Liu, Yihong Gong
arXiv:2603. 11249v4 Announce Type: replace Abstract: Accurate prediction of phase equilibria remains a central challenge in chemical engineering.
By Karim K. Ben Hicham, Moreno Ascani, Jan G. Rittig, Alexander Mitsos
arXiv:2506. 06584v2 Announce Type: replace Abstract: Learning Gaussian Mixture Models (GMMs) is a fundamental problem in statistics and machine learning, with the Expectation-Maximization (EM) algorithm and its popular variant gradient EM being arguably the most widely used algorithms in practice.
By Mo Zhou, Weihang Xu, Maryam Fazel, Simon S. Du
arXiv:2608. 02844v1 Announce Type: cross Abstract: We develop a class of diffusion-based stochastic particle optimisation methods for loss functions with intractable gradients.
By Jiechen Jackie Zhang, O. Deniz Akyildiz
arXiv:2412. 08951v3 Announce Type: replace Abstract: Scalable algorithms of posterior approximation allow Bayesian nonparametrics such as Dirichlet process mixture to scale up to larger dataset at fractional cost.
By Kart-Leong Lim, Xudong Jiang
arXiv:2607. 02292v1 Announce Type: new Abstract: Neural quantum states (NQS) provide a flexible and scalable framework for approximating quantum many-body wavefunctions.
By Juan Agust\'in Duque, Sergio Garc\'ia Heredia, Vinicius Hernandes, Eli\v{s}ka Greplov\'a, Thomas Spriggs, Aaron Courville, Anna Dawid