cs.CV 2510.24795

A Survey on Efficient Vision-Language-Action Models

Proposes resource-efficient VLA design combining model pruning, sparse attention, and training strategies, reducing inference latency to under 50ms.

Zhaoshu Yu, Bo Wang, Pengpeng Zeng et al.

2025-10-28 47