The Embodiment Gap in Robot Foundation Models
Research
Source: arXiv cs.ROPublish time unverified
arXiv:2608.18433v1 Announce Type: new Abstract: Robot foundation models (RFMs), including vision-language-action (VLA) policies, are often discussed through a scaling view: more data, larger models, and broader benchmarks should improve generalization. In robotics, however, a model can generalize while work still remains before it can run on a robot with a particular body. The work required differs across methods and target robots, and those differences affect practical deployment.