arXiv AI
Sep 18

STR-Agent: An LLM-Driven Agent for QoS-Aware Routing in LEO Satellite Networks

STR-Agent is an LLM-driven framework designed for QoS-aware routing in Low Earth Orbit satellite networks. It integrates intent perception, tool-based execution, experience accumulation, and reflection-based policy adaptation to translate natural-language service requests into adaptive routing decisions. In simulations on a Walker-Delta constellation, STR-Agent reduces end-to-end delay by up to 60% compared with DQ-Dijkstra and improves intent-understanding accuracy from 45.4% to 92.45% after fine-tuning, with the Reflection Module providing additional delay reductions.

By Bowen Lu, Mugen Peng, Yaohua Sun, Hongyu Wang, Kerui Guo, Wenjia Xu
arXiv AI
6d ago

Intent2Tc: Automated Intent-to-Traffic Control Translation with Language Models

Intent2Tc is a closed‑loop, language‑model‑driven framework that translates high‑level business traffic‑shaping intents into executable Linux traffic‑control (tc) configurations. It uses an AQM‑based digital twin semantic model, automated metadata extraction, critique‑driven refinement, and Retrieval‑Augmented Generation to improve semantic consistency and configuration reliability. Evaluation on 100 RFC 9315‑compliant intents shows high semantic fidelity and deployment readiness, with Claude Sonnet‑4.6 achieving 0.98 semantic similarity and 0.045 normalized edit distance, while RAG reduces token consumption and latency for compact models.

By Andrea Masini, Sudipta Acharya, Paolo Bellavista, Luca Foschini, Burak Kantarci
arXiv AI
Aug 25

One Request, Multiple Experts: LLM Orchestrates Domain Specific Models via Adaptive Task Routing

The paper introduces ADN‑Agent, an architecture that uses a large language model to orchestrate multiple domain‑specific models (DSMs) for active distribution network (ADN) management. It features adaptive intent recognition, task decomposition, and a unified communication interface for heterogeneous DSMs, along with a pipeline for fine‑tuning small language models on language‑intensive subtasks. Experiments show ADN‑Agent outperforms existing LLM application paradigms in coordinating DSMs for complex ADN operations.

By Xu Yang, Chenhui Lin, Haotian Liu, Qi Wang, Yue Yang, Wenchuan Wu
arXiv AI
Jun 3

vLLM Semantic Router: Signal Driven Decision Routing for Mixture-of-Modality Models

arXiv:2603. 04444v3 Announce Type: replace-cross Abstract: As large language models (LLMs) diversify across modalities, capabilities, and cost profiles, the problem of intelligent request routing -- selecting the right model for each query at inference time -- has become a critical systems challenge.

By Xunzhuo Liu (Steve), Huamin Chen (Steve), Samzong Lu (Steve), Yossi Ovadia (Steve), Guohong Wen (Steve), Hao Wu (Steve), Zhengda Tan (Steve), Jintao Zhang (Steve), Senan Zedan (Steve), Yehudit Kerido (Steve), Liav Weiss (Steve), Haichen Zhang (Steve), Bishen Yu (Steve), Asaad Balum (Steve), Noa Limoy (Steve), Abdallah Samara (Steve), Baofa Fan (Steve), Brent Salisbury (Steve), Ryan Cook (Steve), Zhijie Wang (Steve), Qiping Pan (Steve), Rehan Khan (Steve), Avishek Goswami (Steve), Houston H. Zhang (Steve), Shuyi Wang (Steve), Ziang Tang (Steve), Fang Han (Steve), Zohaib Hassan (Steve), Jianqiao Zheng (Steve), Avinash Changrani (Steve), Xue (Steve), Liu, Bowei He
arXiv Computation and Language
Aug 24

Intent Engine: Natural-Language Intent Translation for Intent-Driven Orchestration in the Compute Continuum

Intent Engine is a natural‑language intent translation architecture that converts user intents into validated Service‑level Objectives (SLOs) for compute‑continuum microservice placement. It combines schema‑constrained extraction, retrieval‑grounded value construction from monitored infrastructure, and validation against supported constraints to produce reliable SLO artifacts. In evaluations on a 716‑record dataset, Intent Engine outperformed prompting baselines and a rule‑based parser, achieving a 0.941 total F1 score with GPT‑4.1 mini and reducing downstream placement failures from 30.8% to 2.1%.

By Koushikur Islam, Rodrigo N. Calheiros