arXiv AI By Zheng Zhang, Liu Liu, Qi Chai, Deheng Ye, Peilin Zhao, Mao Zheng, Hao Wang

Adversarial Closed-Loop Curriculum for Evolving Role-Playing Agents

Read the original on arXiv AI →

The paper introduces AdvRole, an adversarial closed‑loop curriculum for training role‑playing agents with large language models. It alternates between an Actor that learns to role‑play and a Rewriter that edits character profiles and dialogue contexts into hard scenarios, using a performance‑gap reward to target the Actor’s weaknesses. Experiments on English, Chinese, and a new multilingual benchmark demonstrate that AdvRole consistently outperforms baseline methods.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
Sep 15

Learning to Coach for Experiential Learning

arXiv:2609.15851v1 Announce Type: new Abstract: Language models can learn from experience, but raw solution trajectories are often too long and noisy to provide effective guidance. In this work, we p...

By Guanheng Chen, Tianzhu Ye, Li Dong, Xun Wu, Shaohan Huang, Furu Wei