VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving
VLGA introduces a dense 3D geometry expert supervised by LiDAR pointmap reconstruction, achieving state-of-the-art safety and driving scores in autonomous driving benchmarks.
Jin Yao, Dhruva Dixith Kurra, Tom Lampo et al.