arXiv AI

Zonal RL-RRT: Integrated RL-RRT Path Planning with Collision Probability and Zone Connectivity

Zonal RL-RRT is a new path‑planning algorithm that partitions a map into zones using kd‑tree partitioning and employs Value Iteration as a high‑level decision maker. The method achieves a three‑fold improvement in time efficiency over basic sampling methods such as RRT and RRT* in forest‑like maps, and outperforms heuristic‑guided methods like BIT* and Informed RRT* by 1.5× in runtime while maintaining robust success rates across 2D to 6D environments. It also shows on average a 1.5× better performance than learning‑based methods such as NeuralRRT* and MPNetSMP, and has been validated in simulations of a UR10e arm manipulator in MuJoCo.

arXiv AI
Sep 3

DiffuSearch: How Hybrid Trajectory Planning Benefits from Aligned Objectives in Diffusion and Action Space

DiffuSearch is a hybrid trajectory planner for autonomous driving that unifies objectives across both generation and refinement stages. It first uses a guided diffusion model to produce scene-consistent joint trajectories, then refines them with a Monte Carlo Tree Search that shares the same driving goals—collision avoidance, drivable area compliance, comfort, and progress. Experiments on nuPlan and interPlan benchmarks show that this synergy reduces collisions and improves comfort, especially in complex interactive scenarios.

By Steffen Hagedorn, Aron Distelzweig, Alexandru P. Condurache
arXiv AI
Sep 3

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?

The paper investigates Joint-Embedding Predictive World Models (JEPA-WMs), a class of methods that perform planning in a learned representation space rather than raw input space. It systematically studies how model architecture, training objectives, and planning algorithms influence success across simulated and real‑world robotic tasks, and proposes a JEPA-WM variant that surpasses established baselines in navigation and manipulation. The authors provide code, data, and checkpoints for reproducibility.

By Basile Terver, Tsung-Yen Yang, Jean Ponce, Adrien Bardes, Yann LeCun
arXiv AI
Aug 7

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

arXiv:2608. 05588v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires repeatedly planning collision-free paths for agents that continuously receive new goals upon reaching their current ones.

By He Jiang, Jingtian Yan, Yulun Zhang, Yimin Tang, Tanishq Duhan, Rishi Veerapaneni, Guillaume Sartoretti, Jiaoyang Li
arXiv AI
Jul 20

Process Reward Informed Tree Rollout for Effective Multi-Turn RL

arXiv:2607. 15610v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a key approach for training LLM agents, yet popular methods such as GRPO/RLOO rely on multiple independently sampled complete trajectories for advantage estimation.

By Xintong Li, Sha Li, Yuwei Zhang, Changlong Yu, Rongmei Lin, Hongye Jin, Shuyi Guan, Xin Liu, Linwei Li, Qingyu Yin, Jingbo Shang
arXiv AI
Aug 18

LAPF: LLM-Agent-Based Path Finder Using the UAVScenes Dataset

arXiv:2608. 15175v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly deployed for autonomous navigation in complex outdoor environments, where dynamic conditions and mission requirements require intelligent adaptive decision-making.

By Yousef Emami, Mohammadhossein Homaei, Hao Zhou, Miguel Guti\'errez Gait\'an, Atefeh Hajijamali Arani, Rui Zhang