arXiv Machine Learning By Harry Robertshaw, Maxence Boels, Nikola Fischer, Sebastien Ourselin, Christos Bergeles, Alejandro Granados, Thomas C Booth

Progressive Experience Fusion for Multi-Task World Model Control in Endovascular Navigation

Read the original on arXiv Machine Learning →

The paper introduces Progressive Experience Fusion (PEF) for training a multi-task TD-MPC2 controller to navigate endovascular paths across diverse vascular anatomies. PEF outperforms Soft Actor-Critic and base TD-MPC2, achieving 74% success on training anatomies and 90% on held‑out vasculatures. The controller also transfers to an unseen in‑vitro stroke patient, improving path ratio from 63% to 80% after fine‑tuning.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Jul 14

Toward Autonomous Soft Robotic Endovascular Navigation via Imitation Learning

arXiv:2510. 09497v2 Announce Type: replace-cross Abstract: In endovascular surgery, endovascular interventionists push a thin tube called a catheter, guided by a thin wire to a treatment site inside the patient's blood vessels to treat various conditions such as blood clots, aneurysms, and malformations.

By Noah Barnes, Ji Woong Kim, Lingyun Di, Hannah Qu, Anuruddha Bhattacharjee, Miroslaw Janowski, Dheeraj Gandhi, Bailey Felix, Shaopeng Jiang, Olivia Young, Mark Fuge, Ryan D. Sochol, Jeremy D. Brown, Axel Krieger
arXiv Machine Learning
Sep 18

UniExo: Unified Multi-Skill Policies for Musculoskeletal Locomotion and Co-Adaptive Exoskeleton Control

UniExo is a framework that builds a single, multi-skill musculoskeletal human policy by distilling four imitation experts—walking, turning, running, and backward walking—into one network guided by a skill latent. The human policy is fine‑tuned with reinforcement learning on transition sequences, achieving a 94.7% tracking success rate on unseen clips and greater robustness to perturbations. A hip exoskeleton controller is then co‑adapted with this human policy via multi‑agent reinforcement learning, enabling it to assist across four treadmill speeds and a continuous route of all four skills without explicit mode switching.

By Yifei Yuan, Jakob Wolf, Ghaith Androwis, Xianlian Zhou
arXiv AI
Sep 7

Continual Field-Adaptive Models (CFAMs) for Post-Deployment Physical AI

arXiv:2609. 04552v1 Announce Type: cross Abstract: Unattended interactive autonomy - machines that step into danger in place of humans and complete tasks with human tools - remains a missing capability in mission-critical operations.

By Amarjot Singh, Tanmay R. Pancholi, Jainam Kothari, Shrirang Mahajan, Ketan Bansal, Zackory Erickson, Giuseppe Loianno, Alexandre M. Bayen, Jeff Schneider, Vince Nakayama
arXiv AI
Sep 4

FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience

arXiv:2609. 03241v1 Announce Type: cross Abstract: A reasoning model can improve from its own on-policy experience, but this inner loop is fragile: terminal verifiers provide reliable yet sparse supervision, while dense same-model guidance can reinforce false confidence or overconcentrate learning on a narrow solution mode.

By Zixun Huang, Kishan Panaganti, Haitao Mi, Leowei Liang