cs.RO 2510.09459

Failure Prediction at Runtime for Generative Robot Policies

FIPER combines RND and action-chunk entropy for early failure prediction of generative robot policies without failure data, outperforming existing methods.

Ralf Römer, Adrian Kobras, Luca Worbis et al.

2025-10-10 35 citations 27
cs.RO 2510.05430

Active Semantic Perception

Combining large language models with multi-layer scene graphs enables active semantic perception, improving indoor scene understanding speed and accuracy.

Huayi Tang, Pratik Chaudhari

2025-10-07 19
cs.RO 2509.21986

Developing Vision-Language-Action Model from Egocentric Videos

Using EgoScaler to automatically extract 6DoF object trajectories from unlabeled egocentric videos significantly improves VLA pre-training, achieving over 20% success rate gains.

Tomoya Yoshida, Shuhei Kurita, Taichi Nishimura et al.

2025-09-26 41
cs.RO 2509.19142

BiGraspFormer: End-to-End Bimanual Grasp Transformer

BiGraspFormer is an end-to-end transformer framework that directly generates coordinated bimanual grasps from point clouds, achieving high success and efficiency.

Kangmin Kim, Seunghyeok Back, Geonhyup Lee et al.

2025-09-23 58