arXiv Machine Learning By Quanrui Rao, Yong Liu, Xueming Xiao, Yingbo Luo, Kun Wu, Zhenyu Xu, Meibao Yao

RecMorph: Topology-Guided Spatial Recurrence for Generalized Morphology Control

Read the original on arXiv Machine Learning →

RecMorph introduces a topology‑guided spatial recurrent architecture for generalized morphology control, converting a kinematic tree into a sequence that enables joint cross‑limb communication and representation transformation. The design incorporates residual preservation, RMS normalization, and input‑dependent channel modulation to stabilize repeated spatial transformations, achieving linear token complexity. Across five UNIMAL tasks and a four‑platform quadruped setting, RecMorph outperforms existing controllers in training performance, inference throughput, and generalization to unseen bodies with up to 30 limbs, while also demonstrating robust real‑world performance on Go1/Go2 trials.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Aug 24

Graph-Operator World Models for Morphology-Parameter Generalization in Continuous Control

Graph-Operator World Models (GraphOp-WM) are a structured approach to learning world models that generalize across varying morphology parameters in continuous control tasks. The model represents robot bodies and their kinematic relationships as an attributed graph, decomposing each transition into a morphology‑independent local dynamics basis and a morphology‑conditioned structured operator. This operator blends node‑local modulation, kinematic‑tree coupling, and a low‑rank global correction, while architectural design choices encourage the operator to capture static morphology dependence. The framework supports reward, value, and TD‑MPC‑style planning through graph‑level readout and edge‑wise action representations, and is evaluated on controlled MuJoCo parameter splits involving interpolation, extrapolation, and held‑out compositions of link geometry, mass, damping, and actuation in Hopper, Walker2d, and HalfCheetah.

By Xu Yang, Yiqin Yang, Qianchuan Zhao
arXiv AI
Sep 4

BRIDGE: An Open-Source Humanoid Platform via Morphology-Control Co-Design for Physical AI

The paper introduces BRIDGE, an open‑source 88 cm tall humanoid robot designed through a data‑driven morphology‑control co‑design framework that optimizes the robot’s body shape for human‑like movement. A new metric combining kinematic retargeting fidelity and dynamic tracking performance is proposed to evaluate morphological fidelity, and the framework achieves state‑of‑the‑art results compared to existing humanoids such as Bumi, K1, and Toddlerbot. The resulting platform, released with its control policy and supporting materials, demonstrates superior fidelity in capturing human motion, robust balance, and highly dynamic maneuvers.

By Jianren Wang, Letian Qian, Zikai Wang, Weiwei Wu, Junjie Zong, Abhinav Gupta, Deepak Pathak
arXiv AI
Jun 11

Bridging the Morphology Gap: Adapting VLA Models to Dexterous Manipulation via Intent-Conditioned Fine-Tuning

arXiv:2606. 12109v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable zero-shot generalization in robotic manipulation, yet the vast majority of pre-trained pipelines remain strictly confined to low-DoF parallel grippers.

By Chuanke Pang, Junyi Huang, Zhijun Zhao, Yaobing Wang, Kun Xu, Xilun Ding
Hugging Face Trending Papers
Sep 3

BRIDGE: An Open-Source Humanoid Platform via Morphology-Control Co-Design for Physical AI

The paper presents BRIDGE, an open‑source 88 cm tall humanoid robot designed through a data‑driven morphology‑control co‑design framework that aligns robot shape with human‑like movement. It introduces a new metric combining kinematic retargeting fidelity and dynamic tracking performance to evaluate morphological fidelity, achieving state‑of‑the‑art results against baseline humanoids. The released platform, along with its control policy, demonstrates superior human motion capture, robust balance, and dynamic maneuvers, with supporting videos and code available online.