HIL-UMI: Bringing Human-in-the-Loop Post-Training of Vision-Language-Action Models to Universal Manipulation Interface
Problem Adapting large-scale vision-language-action models to specific deployment scenarios is challenging due to the limited coverage of out-of-distribution states and the ineffectiveness of imitation objectives. This paper addresses these issues…
