cs.RO 2608.29896

EMERGE-Policy: A Robot Mind Emerges Beyond a Single Policy

EMERGE-Policy employs a graph-structured multi-agent framework, coordinating specialized sub-agents and skills to achieve system-level robot policies without fine-tuning.

Zhirui Fang, Qingchi Yu, Ziyang Chen et al.

2026-08-31 21
cs.RO 2608.27371

Embodied Scene Rearrangement Planning

Proposes ESRP task with egocentric observations and top-down layouts; introduces ESRP-Bench with 5400+ scene pairs; baseline success only 30.2%.

Canzhi Chen, Zan Wang, Siqi Zhu et al.

2026-08-28 70
cs.RO 2608.25585

RA-VLA: Retrieval-Augmented VLA for Test-Time Adaptation

RA-VLA integrates behavior-aligned retrieval with grounded execution, enabling training-free robotic adaptation with 17.6% success rate improvement on unseen tasks.

Sanghwan Jang, Minjin Jeon, Minsoo Kim et al.

2026-08-26 21
cs.RO 2608.24603

Gripper-aware Vision Language Action Models

Proposes GVLA, integrating multi-gripper encoding and adapter routing, trained on MiGA dataset with 103,000 demonstrations, achieving 7.62% improvement over baselines.

Hanyi Zhang, Zihong Luo, Tianyu Li et al.

2026-08-25 73