arXiv AI By Bo Ai, Henrik I. Christensen, Hao Su

Intelligence Across Embodiments

Read the original on arXiv AI →

The paper discusses how robotic embodiment—sensing, kinematics, dynamics, geometry, actuation, and control—varies across robots and over time, and argues that general embodied intelligence must learn across these differences. It critiques current methods that engineer correspondences for short‑term gains, proposing instead that learning should discover representations that enable transfer across a broader range of embodiments as experience accumulates. The authors advocate for embodiment diversity as a scaling axis, broad learned priors as a complementary ingredient, and evaluations that better characterize embodiment gaps and transfer performance, linking practical cross‑embodiment learning to the scientific pursuit of physical intelligence that adapts with its embodiments.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Aug 20

The Embodiment Gap in Robot Foundation Models

The paper discusses the "embodiment gap" in robot foundation models, highlighting that while models can generalize across tasks, additional work is often needed to deploy them on specific robot bodies. It surveys what components can be reused across different robot embodiments and what must be implemented anew, mapping existing methods along axes of shared structure and adaptation stage. The authors propose a reporting framework to better assess adaptation efforts and identify remaining challenges for cross-embodiment learning.

By Yukiyasu Domae, Keisuke Shirai, Hanbit Oh, Ryoichi Nakajo, Tomohiro Motoda, Koshi Makihara, Masaki Murooka, Takuma Yagi, Yoshiaki Bando, Ryo Hanai
arXiv Machine Learning
Jun 5

Is Diversity All You Need for Scalable Robotic Manipulation?

arXiv:2507. 06219v2 Announce Type: replace-cross Abstract: Data scaling has driven remarkable success in foundation models for Natural Language Processing (NLP) and Computer Vision (CV), yet the principles of effective data scaling in robotic manipulation remain insufficiently understood.

By Modi Shi, Li Chen, Jin Chen, Yuxiang Lu, Chiming Liu, Guanghui Ren, Ping Luo, Di Huang, Maoqing Yao, Hongyang Li
Hugging Face Trending Papers
Jul 13

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence

Artificial general intelligence ultimately requires agents that can reason and act in the physical world. Action models, vision-language-action policies, and world models have advanced this goal, while World Action Models (WAMs) are particularly promising because they connect candidate interventions with predicted consequences.