arXiv AI By Janik Bischoff, Anne Meyer, Uta Mohring, Fabian Dunke, Maximilian Barlang, \"Ozge Nur Subas, Hadi Kutabi, Stefan Nickel, Kai Furmans

Context-Aware Synthesis of Optimization Pipelines for Warehouse Optimization

Read the original on arXiv AI →

arXiv:2606. 26852v1 Announce Type: new Abstract: Order fulfillment in manual picker-to-goods warehouses involves interconnected decisions such as item assignment, order batching, and picker routing.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Sep 4

Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations

The paper presents a graph‑constrained agentic framework that enables large language models to adapt retail supply‑chain decision modules to evolving requirements. It jointly selects intervention routes and admissible module changes, validating candidates against downstream KPIs. Experiments with 100 warehouse requirements and three LLMs show the framework improves end‑to‑end success from 72–76% to 79–83%.

By Lei Zheng, Liping Yang, Zihao Li, Guodong Lyu, Chaik Ming Koh, Chung-Piaw Teo
arXiv AI
Sep 7

Atlas: Optimizing Deployment of Compound AI Workflows on Heterogeneous Clusters

Atlas is a framework that optimizes the deployment of compound AI workflows on heterogeneous clusters by selecting execution plans that satisfy service level objectives (SLOs). It introduces MAP, a Markovian Accuracy Predictor, which estimates configuration accuracy using local conditional accuracy transitions between adjacent workflow stages, avoiding exhaustive end‑to‑end profiling. Atlas formulates plan selection as a mixed‑integer linear program, achieving near‑oracle accuracy while reducing deployment cost by up to 42% and profiling cost by up to 2.6×.

By Milos Gravara, Andrija Stanisic, Stefan Nastic
arXiv Machine Learning
Aug 11

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing

arXiv:2608. 08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries, escalations, and developer wait time are included.

By Srinivasan Manoharan, Junhua Zhao, Fangbo Tu, Haifeng Wu, Jian Wan, Maliah Rajan M, Ashwin Hegde, Mithun Sasidharan, Kalyan Chakravarthi Podamekala