arXiv AI

Intelligence Across Embodiments

The paper discusses how robotic embodiment—sensing, kinematics, dynamics, geometry, actuation, and control—varies across robots and over time, and argues that general embodied intelligence must learn across these differences. It critiques current methods that engineer correspondences for short‑term gains, proposing instead that learning should discover representations that enable transfer across a broader range of embodiments as experience accumulates. The authors advocate for embodiment diversity as a scaling axis, broad learned priors as a complementary ingredient, and evaluations that better characterize embodiment gaps and transfer performance, linking practical cross‑embodiment learning to the scientific pursuit of physical intelligence that adapts with its embodiments.

arXiv Machine Learning
Aug 20

The Embodiment Gap in Robot Foundation Models

The paper discusses the "embodiment gap" in robot foundation models, highlighting that while models can generalize across tasks, additional work is often needed to deploy them on specific robot bodies. It surveys what components can be reused across different robot embodiments and what must be implemented anew, mapping existing methods along axes of shared structure and adaptation stage. The authors propose a reporting framework to better assess adaptation efforts and identify remaining challenges for cross-embodiment learning.

By Yukiyasu Domae, Keisuke Shirai, Hanbit Oh, Ryoichi Nakajo, Tomohiro Motoda, Koshi Makihara, Masaki Murooka, Takuma Yagi, Yoshiaki Bando, Ryo Hanai
arXiv Machine Learning
Jun 5

Is Diversity All You Need for Scalable Robotic Manipulation?

arXiv:2507. 06219v2 Announce Type: replace-cross Abstract: Data scaling has driven remarkable success in foundation models for Natural Language Processing (NLP) and Computer Vision (CV), yet the principles of effective data scaling in robotic manipulation remain insufficiently understood.

By Modi Shi, Li Chen, Jin Chen, Yuxiang Lu, Chiming Liu, Guanghui Ren, Ping Luo, Di Huang, Maoqing Yao, Hongyang Li
Hugging Face Trending Papers
Jul 13

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence

Artificial general intelligence ultimately requires agents that can reason and act in the physical world. Action models, vision-language-action policies, and world models have advanced this goal, while World Action Models (WAMs) are particularly promising because they connect candidate interventions with predicted consequences.

arXiv AI
Sep 21

AtomEgo: Exploring Ego-Robot Integration for Embodied Foundation Model Pretraining

AtomEgo investigates how to integrate large-scale egocentric human interaction data into embodied foundation model pre‑training. The study uses a curated 2,659‑hour corpus and a scalable data pipeline to evaluate three co‑training paradigms across vision‑language‑action and world‑action architectures. Results show that the benefit of egocentric data depends on both its scale and the quality of alignment with robotic embodiment, offering practical guidance for scalable ego‑robot pre‑training.

By Di Wu, Dongchen Zheng, Junhe Sheng, Zhongxing Wei, Songxin Zhang, Zejian Xie, Xiaoquan Sun, Junyang Zheng, Zhuoyang Song, Jiaxing Zhang, Jiayu Chen
arXiv AI
Jun 11

Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models

arXiv:2606. 11324v1 Announce Type: cross Abstract: We introduce Embodied-R1.

By Yifu Yuan, Yaoting Huang, Xianze Yao, Yutong Li, Shuoheng Zhang, Linqi Han, Pengyi Li, Jiangeng Sun, Wenting Jia, Zhao Zhang, Yuhao Liu, Ruihao Liao, Yucheng Hu, Qiyu Wu, Yuxiao Li, Zibin Dong, Fei Ni, Yan Zheng, Shuyang Gu, Yi Ma, Hongyao Tang, Han Hu, Jianye Hao