具身智能观察

可指令智能体的规划与控制解耦

原标题:Decoupling Planning and Control for Instructable Agents

产业动态AI 72

来源:arXiv cs.RO发布时间待核实

arXiv:2608.26788v1 Announce Type: cross Abstract: Recent work shows that pre-trained, instruction-tuned vision-language models (VLMs) perform well at mapping from instructions and observations to high-level plans, but struggle to realize such plans as reliable low-latency action sequences in unfamiliar environments. At the same time, world-model controllers excel at fast observation-to-action control, but lack open-ended task guidance.

可指令智能体的规划与控制解耦 | 具身智能观察