arXiv AI By Zeyu Ren, Ling Yue, Ran Li, Yishu Wang, Shengxiang Xu, Hanmo Liu, Shaowu Pan, Shimin Di

FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills

Read the original on arXiv AI →

arXiv:2607. 21596v2 Announce Type: replace Abstract: Large language model agents can adapt to complex tasks by constructing workflows at inference time, but procedures discovered in one episode are usually discarded after execution.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
4d ago

It Takes Workflows to Evolve Better Workflows

The paper "It Takes Workflows to Evolve Better Workflows" introduces FloWright, a method that uses a hierarchical, structure‑aware reward system to allow one or more roles in a multi‑agent workflow to self‑evolve without extra models or data. It also proposes DataWright, an adaptive data hardening technique that transforms existing datasets into more challenging workflow‑level tasks. Experiments on document, slide, chart, code, math, and finance tasks show that small open models trained with FloWright can improve performance by up to +7.41%, with co‑evolving multiple roles yielding the largest gains. "whyItMatters":"The work demonstrates that optimizing beyond the workflow generator—by enabling multiple agents to co‑evolve—can substantially enhance the effectiveness of multi‑agent workflows for complex real‑world tasks."

By Xuehang Guo, Haoyu Wang, Haifeng Chen, Yangyi Chen, Zhenhailong Wang, Qingyun Wang
arXiv AI
Jul 7

EvoAgentBench: Benchmarking Agent Self-Evolution via Ability Transfer

arXiv:2607. 05202v1 Announce Type: new Abstract: Agent self-evolution in long-horizon LLM systems is largely procedural: useful experience is not merely stored information, but reusable procedures for searching, debugging, and verification.

By Xingze Gao, Chuanrui Hu, Hongda Chen, Pengfei Yao, Zhao Wang, Yi Bai, Zhengwei Wu, Yunyun Han, Xiaofeng Cong, Jie Gui, Yafeng Deng, Teng Li
arXiv AI
Sep 17

Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents

The paper introduces EvoSkill-GUI, a training‑free framework that enables GUI agents to evolve their skills during deployment. Each skill is packaged with metadata, executable plans, and recovery rules, and the system follows a reflect‑revise‑reuse loop where the agent instantly revises skills based on execution feedback. Experiments on MobileWorld, AndroidWorld, and OSWorld show consistent performance gains up to +16.2% without any additional training.

By Bofan Chen, Boxuan Zhang, Fei Tang, Zhengxi Lu, Yong Du, Tongbo Chen, Weiming Lu, Jun Xiao, Yueting Zhuang, Yongliang Shen