CORTEXA
← Browse

Shaopeng Zhai

1 paper indexed

arxivcs.ROcs.AI2026-07-08

HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation

Shaopeng Zhai, Qi Zhang, Tianyi Zhang, Haoran Zhang, Fuxian Huang, Zhanhui Lin, et al.

When adapting Vision Language Action (VLA) models to downstream tasks, multiple rounds of post-training are often required to progressively address policy weaknesses. In this report, we focus on maximizing human efficiency during this iterative process, measured by policy improve…

View free PDFSource page