arXiv:2608. 14443v1 Announce Type: cross Abstract: Neural Architecture Search (NAS) is naturally formulated as a bilevel optimization problem, where the upper-level optimizes the architecture using validation performance and the lower-level trains network parameters using training loss.
By Abhishek Shukla, Ankur Sinha, Faiz Hamid
arXiv:2608. 14472v1 Announce Type: cross Abstract: Neural Architecture Search (NAS) aims to automate neural network architecture design, reducing reliance on human expertise.
By Abhishek Shukla, Ankur Sinha, Faiz Hamid
arXiv:2609.22329v1 Announce Type: cross
Abstract: Black-Box Optimization (BBO) is often applied in several engineering fields and can utilize an advancement of numerical measure- ments and simulation...
By Md Khadimul Islam Zim (Czech Academy of Sciences, Institute of Computer Science, Prague, Czech Republic), Martin Hole\v{n}a (Czech Academy of Sciences, Institute of Computer Science, Prague, Czech Republic)
arXiv:2608. 12704v1 Announce Type: cross Abstract: Multi-objective bilevel optimization has wide applications in the AI area such as automated learning and multi-task meta-learning.
By Yicong Jiang, Feihu Huang
arXiv:2606. 00862v1 Announce Type: cross Abstract: Surrogate-assisted evolutionary algorithms (SAEAs) have been widely used for expensive black-box optimization problems.
By Xiao Jin, Yongxiong Wang, Haobo Liu, Yudong Du, Yukun Du
arXiv:2607. 06772v1 Announce Type: new Abstract: Learned optimization aims to improve upon hand-designed optimizers (e.
By Xiaolong Huang, Benjamin Th\'erien, James Harrison, Eugene Belilovsky
The paper introduces BOTH, a method that differentiates topology optimization (TO) itself to compute hypergradients for tuning hyperparameters alongside the primary design optimization. By evaluating only one or two TO steps, the approach provides sufficient information and scales to thousands of hyperparameters with a cost comparable to a few standard TO runs. Experiments on stress‑constrained and compliance problems, including a neural‑parameterized density field, demonstrate the effectiveness of this joint optimization strategy.
By Suryanarayanan Manoj Sanu, Miguel Anibal Bessa, Alejandro Marcos Arag\'on
arXiv:2609.35827v1 Announce Type: cross
Abstract: The quality of a detector design is ultimately determined by the quality of the inference it enables, that is, by the accuracy with which the quantit...
By Maxim Borisyak, Nikita Gladin, Andrey Ustyuzhanin
arXiv:2607. 04033v1 Announce Type: cross Abstract: Optimizer selection for large-scale model training has become a system-level design decision constrained jointly by compute, memory, tuning budget, and task diversity, yet the landscape of over one hundred methods remains fragmented.
By Siyuan Li, Jiabao Pan, Yumou Liu, Zhuoli Ouyang, Xin Jin, Xinglong Xu, Jingxuan Wei, Shengye Pang, Jintao Che, Xuanhe Zhou, Conghui He, Cheng Tan
Learned optimization aims to improve upon hand-designed optimizers (e. g.
arXiv:2405. 04376v4 Announce Type: replace Abstract: Hyperparameter tuning, particularly the selection of an appropriate learning rate in adaptive gradient training methods, remains a challenge.
By Yijiang Pang, Shuyang Yu, Bao Hoang, Jiayu Zhou
arXiv:2606. 14970v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) has become a central application of modern optimization, enabling pretrained models to adapt to diverse downstream tasks and domain-specific data.
By Dmitriy Bystrov, Daniil Medyakov, Dmitry Bylinkin, Aleksandr Beznosikov