arXiv Machine Learning By Rayhan Patel, Shabaz Patel

EvoRank: LLM-Guided Evolution of Multi-Objective Learning-to-Rank Pipelines

Read the original on arXiv Machine Learning →

EvoRank is an open autonomous ranking engineer that uses an LLM-guided evolutionary loop to automatically design complete Learning-to-Rank pipelines—including features, models, losses, and ensembles—for multi-objective e-commerce search. On the Expedia ICDM 2013 dataset, EvoRank converged within 50 iterations on interpretable pipelines that outperform an Optuna-tuned LambdaMART on 60k held-out queries and rank in the top 6 % of the original competition. The authors also introduce a headroom gate that predicts whether the evolutionary loop will be worthwhile before any LLM computation, and they release the system, auditing tools, and a catalog of failure modes to help teams apply the method to their own ranking stacks.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 18

Evolution or Illusion? Rethinking Evaluation in LLM Evolutionary Search

The paper critiques the common practice of evaluating large‑language‑model (LLM) evolutionary search methods using a single seed and fixed iteration budget, arguing that this approach is insufficient. By testing three search strategies across five optimization tasks and varying both the number of seeds (width) and iterations (depth), the authors find that optimal budget allocation depends on the strategy, task, and total budget, and that strategy rankings shift with different budgets. They propose a measurement protocol that maps the seeds‑by‑iterations frontier and offers practical guidance for researchers.

By Tal Oved, Roi Pony, Oshri Naparstek, Udi Barzelay
arXiv Machine Learning
Aug 4

When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design

arXiv:2608. 01378v1 Announce Type: new Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensive evaluation - an experiment, a first-principles simulation, or a full training run.

By Shuangxiu (Max), Ma (Zachary), Wenhe (Zachary), Zhao
arXiv AI
Aug 25

Enrich-Retrieve-Rank: Scaling Capability Discovery Beyond In-Context Routing

The paper introduces Enrich‑Retrieve‑Rank, a scalable method for discovering capabilities in large agent ecosystems. It replaces in‑context routing with an offline enrichment step that converts sparse metadata into searchable profiles, followed by an online retrieve‑then‑rank pipeline that returns a ranked shortlist without invoking candidates. Experiments show that as the number of capabilities grows from 10 to 7,278, the new approach maintains higher top‑1 accuracy and reduces cost by 70× compared to full‑context baselines.

By Nazib Sorathiya, Daniel Zhang, Bardiya Akhbari