cs.RO 2409.14562

DROP: Dexterous Reorientation via Online Planning

DROP employs online sampling predictive control with vision-based pose estimation to reorient objects, achieving performance comparable to RL methods without extensive training.

Albert H. Li, Preston Culbertson, Vince Kurtz et al.

2024-09-23 26
cs.RO 2407.12957

R+X: Retrieval and Execution from Everyday Human Videos

R+X leverages vision-language models for retrieval and in-context imitation learning, enabling robots to learn from unlabelled human videos without training.

Georgios Papagiannis, Norman Di Palo, Pietro Vitiello et al.

2024-07-18 55 citations 109