arXiv AI

MOSAIC: Adversarial Co-evolution of Specialist Heuristics and Problem Instances for LLM-based Automated Heuristic Design

arXiv:2608. 07544v1 Announce Type: cross Abstract: Automated heuristic design (AHD) with large language models (LLMs) has produced strong heuristics for combinatorial optimization problems (COPs).

arXiv AI
Sep 10

An Evolutionary Framework for Automatic Optimization Benchmark Generation via Large Language Models

The paper introduces LLM-EBG, an evolutionary framework that uses a large language model as a generative operator to automatically create optimization benchmarks. By generating unconstrained single-objective continuous minimization problems expressed as mathematical formulas, the framework can produce benchmarks that consistently favor a target algorithm over a comparison algorithm in over 80% of trials. Landscape analysis shows that these generated problems exhibit distinct geometric traits, such as sensitivity to variable scaling, reflecting the search behaviors of different optimization methods.

By Yuhiro Ono, Tomohiro Harada, Yukiya Miura
arXiv AI
Sep 3

Adaptive Graph-of-Islands Evolution for Automatic Feature Engineering with LLMs

The paper introduces TOPOFE, a framework that treats automatic feature engineering for tabular data as a graph-structured multi-island evolutionary search. Each island explores a semantically coherent family of transformations using LLM-guided mutation and crossover, while a Prompt Adaptation Memory steers proposals based on accept/reject feedback. TOPOFE dynamically learns a directed topology graph to coordinate cross-island transfer, enabling the discovery of compositional feature programs that outperform state‑of‑the‑art methods on 29 datasets and produce lower redundancy and higher representational coverage.

By Sha Li, Naren Ramakrishnan
arXiv AI
Jun 2

BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution

arXiv:2606. 01286v1 Announce Type: cross Abstract: The rapid progress of frontier large language models has led to widespread benchmark saturation, limiting the ability of existing datasets to differentiate model capabilities or provide useful training signal.

By Yangzhen Wu, Aaron J. Li, Wenjie Ma, Li Cao, Ziheng Zhou, Mert Cemri, Shu Liu, Yuran Xiu, Chenxiao Yan, Haikun Zhao, Bin Yu, Ion Stoica, Dawn Song