具身智能观察

Learning to Act While Waiting: RL Finetuning of Generalist Robot Policies Under Inference Latency

产业动态

来源:arXiv cs.RO发布时间待核实

arXiv:2608.23831v2 Announce Type: replace Abstract: While reinforcement learning (RL) allows generalist robot policies to continually improve during deployment, the large model size of modern generalist policies, such as VLAs, poses a fundamental obstacle to effective RL improvement.