arXiv AI
Aug 26

MARS: Multi-Specialist LLM Relay System for Competitive Programming

MARS (Multi-Agent Relay of Specialized LLMs) is a prompt-only framework that assigns specialized LLM agents—each focused on a particular algorithmic domain such as dynamic programming, graphs, or geometry—to collaboratively solve competitive programming problems. Retrieval-augmented generation selects a small team of relevant specialists for each problem, and the agents iteratively refine a C++17 solution through sandboxed testing, passing structured packets between them until a final infrastructure-fixer normalizes the code. On the CodeContests benchmark, MARS achieves a pass rate of 0.624 with Gemma 4, improving over direct prompting by 14.4 percentage points while reducing wall‑clock cost and token‑spend variance compared to CodeSIM.

By Andrei Mikhailov, Mikhail Burtsev, Alsu Sagirova
arXiv AI
Jul 1

ClawArena-Team: Benchmarking Subagent Orchestration and Dynamic Workflows in Language-Model Agents

arXiv:2606. 31174v1 Announce Type: new Abstract: Production large language-model (LLM) agents are increasingly deployed not as lone problem-solvers but as managers: a main model creates specialized subagents, delegates work, and orchestrates their parallel, asynchronous returns through dynamic workflows.

By Kaiwen Xiong, Haonian Ji, Shi Qiu, Zeyu Zheng, Cihang Xie, Xinyu Ye, Huaxiu Yao