arXiv AI By Lei Zheng, Liping Yang, Zihao Li, Guodong Lyu, Chaik Ming Koh, Chung-Piaw Teo

Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations

Read the original on arXiv AI →

The paper presents a graph‑constrained agentic framework that enables large language models to adapt retail supply‑chain decision modules to evolving requirements. It jointly selects intervention routes and admissible module changes, validating candidates against downstream KPIs. Experiments with 100 warehouse requirements and three LLMs show the framework improves end‑to‑end success from 72–76% to 79–83%.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Aug 18

Agent Gym: A Framework for Continuous Evaluation and Evolution of LLM Agents Through Human-in-the-Loop Feedback

arXiv:2608. 15591v1 Announce Type: new Abstract: Large Language Model (LLM) agents deployed in production environments face a fundamental tension: the agent's behavior is frozen at deployment time, while the business rules and edge cases it must handle continue to evolve.

By Pouya Ghiasnezhad Omran, Michael Zimmermann, Duncan Cambridge, Ashmita Kapoor, Tanya Dixit
arXiv AI
Sep 10

Learning to Configure Agentic AI Systems

The paper introduces ARC, a lightweight hierarchical policy that learns to configure LLM‑based agent systems on a per‑query basis by treating each configuration as a temporally extended option in a semi‑Markov decision process. Unlike fixed templates or hand‑tuned heuristics, ARC dynamically selects workflows, tools, token budgets, and prompts tailored to the difficulty of each query. Experiments on reasoning, tool‑use, and agentic benchmarks show that ARC outperforms budget‑matched tool‑augmented LLMs, boosting reasoning accuracy by 31.3%, tool‑use accuracy by 13.95%, and doubling success on the τ‑Bench Airline Pass task from 9.0% to 18.0%.

By Aditya Taparia, Som Sagar, Ransalu Senanayake
arXiv AI
Jul 21

Agentic ERP: Multi-Agent Large Language Model Architecture for Autonomous Enterprise Resource Planning

arXiv:2607. 17331v1 Announce Type: new Abstract: Enterprise Resource Planning (ERP) systems record transactions reliably but still delegate almost all operational decision-making to human specialists, because classical rule-based automation cannot reason about exceptions and monolithic AI assistants degrade when asked to coordinate across functional boundaries.

By Zhihao Liu, Tianyu Wang, Xi Vincent Wang, Lihui Wang