Embodied Intelligence Observer

PredVLA: A Sub-Million-Parameter Predictive-Coding Policy for Robot Manipulation

Industry

Source: arXiv cs.ROPublish time unverified

arXiv:2608.26673v1 Announce Type: new Abstract: Large pretrained vision-language-action models dominate modern robot-manipulation benchmarks, but it remains unclear how much model scale is necessary for strong language-conditioned control, or whether fundamentally different control architectures can remain competitive at much smaller parameter budgets.

PredVLA: A Sub-Million-Parameter Predictive-Coding Policy for Robot Manipulation | Embodied Intelligence Observer