arXiv AI By Ishaan Bhola, Adithyan Krishnan, Mukunda NS

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First

Read the original on arXiv AI →

arXiv:2608. 04804v1 Announce Type: cross Abstract: Frontier language models can resolve repository-level software issues, but each attempt is expensive, and existing routers select a model from the issue text alone.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
3d ago

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

arXiv:2605.18859v3 Announce Type: replace-cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single u...

By Pei Yang, Wanyi Chen, Tongyun Yang, Pengbin Feng, Jiarong Xing, Wentao Guo, Yuhang Yao, Yuhang Han, Hanchen Li, Xu Wang, Zeyu Wang, Jie Xiao, Anjie Yang, Liang Tian, Lynn Ai, Eric Yang, Tianyu Shi
arXiv AI
Aug 26

MARS: Multi-Specialist LLM Relay System for Competitive Programming

MARS (Multi-Agent Relay of Specialized LLMs) is a prompt-only framework that assigns specialized LLM agents—each focused on a particular algorithmic domain such as dynamic programming, graphs, or geometry—to collaboratively solve competitive programming problems. Retrieval-augmented generation selects a small team of relevant specialists for each problem, and the agents iteratively refine a C++17 solution through sandboxed testing, passing structured packets between them until a final infrastructure-fixer normalizes the code. On the CodeContests benchmark, MARS achieves a pass rate of 0.624 with Gemma 4, improving over direct prompting by 14.4 percentage points while reducing wall‑clock cost and token‑spend variance compared to CodeSIM.

By Andrei Mikhailov, Mikhail Burtsev, Alsu Sagirova