Embodied Intelligence Observer

The Embodiment Gap in Robot Foundation Models

Research

Source: arXiv cs.ROPublish time unverified

arXiv:2608.18433v1 Announce Type: new Abstract: Robot foundation models (RFMs), including vision-language-action (VLA) policies, are often discussed through a scaling view: more data, larger models, and broader benchmarks should improve generalization. In robotics, however, a model can generalize while work still remains before it can run on a robot with a particular body. The work required differs across methods and target robots, and those differences affect practical deployment.

The Embodiment Gap in Robot Foundation Models | Embodied Intelligence Observer