D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models
D-VLA framework uses 'Plane Decoupling' and a four-thread asynchronous pipeline to enhance VLA model sampling efficiency and throughput.
Yucheng Guo, Yongjian Guo, Zhong Guan et al.