arXiv:2606. 04167v1 Announce Type: cross Abstract: We tackle the Metro Network Expansion Problem (MNEP), a subset of the Transport Network Design Problem (TNDP), which focuses on expanding metro systems to satisfy travel demand.
By Dimitris Michailidis, Sennay Ghebreab, Fernando P. Santos
We’ve developed a hierarchical reinforcement learning algorithm that learns high-level actions useful for solving a range of tasks, allowing fast solving of tasks requiring thousands of timesteps. Our algorithm, when applied to a set of navigation problems, discovers a set of high-level actions for walking and crawling in different directions, which enables the agent to master new navigation tasks quickly.
An expert in behavioral science and transportation, Zhao combines these studies with AI and public policy to address some of the most urgent challenges facing cities.
By Maria Iacobo | School of Architecture and Planning
We’re launching a transfer learning contest that measures a reinforcement learning algorithm’s ability to generalize from previous experience.
arXiv:2604. 17456v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown strong capabilities in long-horizon reasoning, tool use, and decision-making in digital environments, yet extending them to physically grounded systems remains challenging.
By Siqi Lai, Pan Zhang, Yuping Zhou, Jindong Han, Yansong Ning, Hao Liu
The paper introduces an agent‑based modeling framework that integrates a physical infrastructure layer, a V2X connectivity layer, and a decision layer using reinforcement learning and multi‑agent reinforcement learning to simulate smart freight corridors. Three scenarios—Baseline, Assisted, and Cognitive—are evaluated on throughput, congestion, energy, emissions, and robustness, with the Cognitive scenario outperforming the baseline in throughput and congestion, and the Assisted scenario achieving energy savings via platooning. Sensitivity analysis shows that the smart corridor’s throughput advantage grows under high demand and that MARL coordination better utilizes fixed charging capacity than rule‑based methods.
By Madelaine Martinez-Ferguson, Chun Wang, Mustafa Can Camur, Xueping Li