MILER is an end‑to‑end reinforcement learning framework that achieves zero‑shot sim‑to‑real transfer for autonomous driving in unstructured environments. It uses a custom semantic mid‑level representation (MLR) simulator for offline training, and during deployment it processes real camera and LiDAR data with BEVFusion to produce a compatible bird’s‑eye‑view representation. The policy’s actions are applied via a trajectory‑alignment strategy, allowing the system to drive 17.3 km on a 3.0 km test track without human intervention, all running on a Jetson AGX Orin.
By Thomas Steinecker, Denis Trescher, Alexander Bienemann, Thorsten Luettel, Mirko Maehlisch
arXiv:2607. 11349v1 Announce Type: cross Abstract: Dual-source trolleybuses alternate between overhead catenary supply and on-board battery operation, creating energy-use patterns driven by route attributes, high-frequency trajectories, and hourly weather.
By Wentao Zeng (School of Management, Foshan University, Foshan, China a School of Management, Foshan University, Foshan, China, School of Mechanical and Electrical Engineering and Automation, Foshan University, Foshan, China), Zijian Huang (School of Artificial Intelligence, South China Normal University, Guangzhou, China), Yiming Bie (School of Transportation, Jilin University, Changchun, China), Jiabin Wu (School of Management, Foshan University, Foshan, China a School of Management, Foshan University, Foshan, China), Jun Gong (Department of Civil Engineering, The University of Hong Kong, Hong Kong, China)
arXiv:2609.39868v1 Announce Type: new
Abstract: Exploratory reinforcement learning (RL) on an operating bus fleet is impractical,while policies trained only from historical data cannot acquire new ex...
By Yifan Zhang, Qifan Zhang, Liang Zheng
arXiv:2509. 21842v2 Announce Type: replace Abstract: Travel planning (TP) agent has recently worked as an emerging building block to interact with external tools/resources for travel itinerary generation, ensuring an enjoyable user experience.
By Yansong Ning, Rui Liu, Jun Wang, Kai Chen, Wei Li, Jun Fang, Kan Zheng, Naiqiang Tan, Hao Liu
arXiv:2609.21945v1 Announce Type: new
Abstract: Urban transportation networks present complex optimization challenges spanning calibration of high-fidelity simulators and real-time operational contro...
By Adewumi Augustine Adepitan, Christopher J. Haruna, Oluwasegun Adegoke, Ayooluwatomiwa Ajiboye, Oluwatobi Oluwasakin
arXiv:2604. 03497v2 Announce Type: replace-cross Abstract: Vision-language-model (VLM)-guided reinforcement learning (RL) has recently attracted significant attention for it, replacing brittle hand-crafted rewards with semantically grounded signals; however, deploying such simulation-trained policies on real vehicles remains a fundamental challenge, because they rely on simulator-native observations and simulator-coupled action semantics with no counterpart on physical hardware.
By Zilin Huang, Zhengyang Wan, Zihao Sheng, Boyue Wang, Junwei You, Sikai Chen