arxivcs.ROcs.AI2026-07-08
HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation
Shaopeng Zhai, Qi Zhang, Tianyi Zhang, Haoran Zhang, Fuxian Huang, Zhanhui Lin, et al.
When adapting Vision Language Action (VLA) models to downstream tasks, multiple rounds of post-training are often required to progressively address policy weaknesses. In this report, we focus on maximizing human efficiency during this iterative process, measured by policy improve…