arXiv AI By Jianming Chen, Xuanbin Ye, Yawen Wang, Junjie Wang, Qing Wang, Fanjiang XU

VCE-Skill: Enhancing Skill Self-Evolution with Version-Change Experience

Read the original on arXiv AI →

arXiv:2608. 16544v1 Announce Type: cross Abstract: Agents increasingly rely on reusable skills to encode task knowledge, tool-use procedures, and validation rules.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Sep 25

A Wrong Turn Does Not Ruin the Journey: Deviation-Guided Skill Self-Evolution for LLM Agents

The paper introduces SkillPivot, a framework that guides large language model agents to evolve their skills by pinpointing the exact moment a useful problem‑solving sequence turns into an erroneous suffix. SkillPivot uses execution validity, goal progress, and action diversity to detect this deviation point, then employs a stronger teacher to generate a successful alternative from the same prefix. By contrasting the failed and successful suffixes, the method produces localized, compact skill updates that preserve existing effective guidance, outperforming other skill‑evolution techniques on benchmarks such as ToolQA, LogicBench, and WildClawBench.

By Yichun Feng, Jiawei Wang, Haozhe Sun
arXiv AI
6d ago

SkillEvoReg: Regularizing Agent Skill Evolution Against Overfitting

SkillEvoReg is a regularization framework designed to mitigate overfitting in language-model agents that evolve reusable external skills. It combines training-time skill dropout, complexity-aware local regularization, and causal counterexample validation to control skill-state growth and detect regressions. Applied across SkillOpt, SkillEvolBench, and ContinualSkillBench, it preserves downstream performance while improving transfer and later-stage evolution outcomes.

By Guanyu Nie, Fangzhou Zhu, Shixiong Kai, Xiongwei Han, Tao Zhong, Mingxuan Yuan
Hugging Face Trending Papers
Sep 24

A Wrong Turn Does Not Ruin the Journey: Deviation-Guided Skill Self-Evolution for LLM Agents

The paper introduces SkillPivot, a framework that guides large language model agents to improve their natural-language skills by focusing on the point where a successful solution path deviates into an error. SkillPivot identifies this transition using execution validity, goal progress, and action diversity, then employs a stronger teacher to generate a successful alternative from the same prefix. By contrasting the failed and successful suffixes, the method produces localized, compact skill updates that preserve existing effective guidance and outperform other skill-evolution approaches on multiple benchmarks.