具身智能观察

Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization

技术动态

来源:arXiv cs.RO发布时间待核实

arXiv:2608.26103v1 Announce Type: new Abstract: Zero-shot cross-task generalization, where a policy must execute manipulation tasks never seen during training, remains a central challenge in robot learning. In large language models, a novel task can be performed simply by specifying it in the context, without any parameter update. This form of in-context learning (ICL) turns generalization into a problem of task specification.

多源报道2

查看事件全景 →
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization | 具身智能观察