具身智能观察

Mamba 选择性状态空间改进 SmolVLA 精度-复杂度权衡

原标题:Mamba-based Selective State Space Modeling Improves the Accuracy-Complexity Tradeoff of SmolVLA Vision-Language-Action Experts

产业动态AI 66

来源:arXiv cs.RO发布时间待核实

arXiv:2608.21407v1 Announce Type: new Abstract: Vision-language-action (VLA) models face a crucial tradeoff between their task success rate and the policy-call frequency. Executing a single action per inference ($N=1$) enables accurate robot control but comes at the cost of huge compute time overheads, making real-time implementation infeasible.