ORPA: Online Residual Policy Adaptation for Robot Manipulation Control with Human Feedback

arXiv:2608.17323v1 Announce Type: new Abstract: Robotic manipulation policies trained via imitation learning, such as Action Chunking with Transformers (ACT), can achieve strong performance under ideal conditions but often remain sensitive to small execution errors and distribution shifts. Correcting these failures typically requires dataset aggregation and full-policy retraining, which is computationally expensive and unsuitable for real-time deployment. In this work, we propose Online Residual...

arXiv cs.RO ·Muhammad A. Muttaqien, Tomohiro Motoda, Ryo Hanai, Yukiyasu Domae ·
compartilhar: