From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for Embodied Intelligence

arXiv:2607.26903v1 Announce Type: new Abstract: The key bottleneck in embodied AI is not model architecture but data. Although billions of human manipulation videos exist online, robots cannot directly learn from them due to the embodiment gap between human morphology and robot hardware. We introduce Pegasus, a low-resource framework that bridges this gap by translating human demonstrations into robot-learnable data through structured knowledge transfer. Instead of relying on raw video prompts, ...

arXiv cs.AI ·Jia Luo ·
compartilhar: