GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.RO 2608.13396

Capstan-driven Continuum Surgical Robot: Design, Modeling, and Perception

Proposes actuation-perception co-design with micro-deformation sensing and short-thick-beam modeling, achieving real-time shape and force perception in compact capstan-driven continuum robots.

Gang Zhang, Yufu Qiu, Junyan Yan et al.

2026-08-13 100
cs.RO 2608.11731

ContactIPM: A Structure-Exploiting Interior-Point Solver for Contact-Implicit Trajectory Optimization

ContactIPM combines structure-exploiting interior-point method with stagewise elastic relaxation, achieving 2-8x speedup in contact-implicit trajectory optimization.

Yucheng Chen

2026-08-12 87
cs.RO 2608.11521

Keep the Future, Drop the Rollout: RIFT for World Action Models

RIFT constructs full future K/V cache in one pass using learned anticipation tokens, achieving 98.8% success and reducing latency by up to 89%.

Chushan Zhang, Jinguang Tong, Xuesong Li et al.

2026-08-12 45
cs.RO 2608.11174

VIScore: Diagnosing Planning-Relevant Quality in Latent World Models

Proposes VIScore, integrating veracity, influence, and sobriety, to evaluate planning-relevant quality in latent world models, with correlation exceeding 0.75.

Haiyu Wu, Randall Balestriero, Morgan Levine

2026-08-12 135
cs.RO 2608.09892

XPolicyLab: A Unified Standard and Open Ecosystem for Robot Policy Evaluation and Deployment

XPolicyLab introduces a unified standard for robot policy interfaces, reducing integration complexity from O(NM) to O(N+M), enabling seamless evaluation of 42 diverse policies across platforms.

XPolicyLab Community, Tianxing Chen, Yue Chen et al.

2026-08-11 91
cs.RO 2608.08053

PhysX-CoT: Structured Physical Reasoning from a Single Image to Simulation-Ready 3D Assets

PhysX-CoT uses structured physical reasoning to generate simulation-ready 3D assets from a single image, outperforming existing baselines.

Jie Huang, Xiaohe Li, Jiahao Li et al.

2026-08-08 30
cs.RO 2608.07361

Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model

This study introduces depth-wise probing and pruning of the planning token in a driving VLA model, revealing early linear decodability of semantic intent and enabling 1.33× speedup with minimal performance loss.

Harisankar Babu, Benjamin Coors, Christopher Lang et al.

2026-08-08 149
cs.RO 2608.06994

Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models

PILOT framework uses representational deduction with Motion CoT to decouple high-level intention from low-level trajectories, achieving 97.9% success on LIBERO.

Xiangkai Ma, Yue Ma, Junjie Wang et al.

2026-08-07 42
cs.RO 2608.21400

Active Interaction-Aware Model Predictive Path Integral via Ego-Conditioned Generative Predictions

Proposed an Active Interaction-Aware Path Integral method using Ego-Conditioned Generative Predictions to enhance safety and efficiency in dense traffic scenarios.

Khaled A. Mustafa, Mohamed-Khalil Bouzidi, Christian Schlauch et al.

2026-08-07 3
cs.RO 2608.06208

ErgoSurf: Ergodic Control for the Coverage of Unknown Surfaces

Proposes ErgoSurf combining online GPIS surface reconstruction with ergodic control for unknown surfaces, achieving near-ground-truth accuracy.

Stefan Schneyer, Timo Bachmann, Maged Iskandar et al.

2026-08-07 94
cs.RO 2608.06170

Prior-SG: Task and Prior Driven Region Segmentation for Scene Graphs in Arbitrarily-Structured Environments

Prior-SG formulates scene graph generation as a probabilistic alignment problem, integrating multi-scale feature fusion and graph optimization to achieve robust semantic region segmentation in complex environments.

Giorgio Tonetti, Laurent Kneip, Abel Gawel et al.

2026-08-06 99
cs.RO 2608.05588

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

Introduced SJRL method, significantly enhancing multi-agent pathfinding performance, especially on high-density maps.

He Jiang, Jingtian Yan, Yulun Zhang et al.

2026-08-06 5
cs.RO 2608.05042

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

BridgeVLA++ integrates multi-view projection, explicit spatio-temporal memory, and pre-trained VLMs for efficient, robust 3D robot manipulation.

Peiyan Li, Yuze Zhu, Yixiang Chen et al.

2026-08-06 126
cs.RO 2608.03820

Designing Social Robots for Inclusive Child Wellbeing Assessment: Insights from Communities Supporting Developmental Language Disorder and Forced Migration

Developed a multimodal robot interaction framework supporting inclusive wellbeing assessment for children with DLD and migration backgrounds, with ethical design principles.

Fethiye Irmak Dogan, Yue Lou, Alva Markelius et al.

2026-08-04 48
cs.RO 2608.03231

Structure-Aware Robust Fine-Tuning: Defending Vision-Language-Action Robots Against Physical Attention Hijacking

SARF fine-tunes only the visual encoder with feature anchoring and attention correction, reducing failure rate from 100% to 14.2%-56.8% against physical attention hijacking.

Jinquan Zhang, Dongfu Yin, Run Yang et al.

2026-08-04 53
cs.RO 2608.01834

Teleopit: A Full-Embodiment Humanoid Teleoperation System

Teleopit employs VR-based full-body, hand, and viewpoint mapping, achieving 95% task success on humanoid robots.

Bingqian Wu, Zicheng Xu, Xianghui Fan et al.

2026-08-03 52
cs.RO 2608.01600

Perception-and-action system for humanoid robot task execution in construction

Proposes Humanoid-PoseNet and Humanoid-ActionNet for learning construction tasks from worker demonstrations, achieving 82.45mm MPJPE accuracy.

Yanxi Liu, Yizhi Liu

2026-08-03 46
cs.RO 2608.01452

DynamicManip: Enabling Dynamic Manipulation from a Single Static Demonstration

DynamicManip synthesizes diverse dynamic demonstrations from a single static example, boosting data efficiency and real-time response in robot manipulation.

Haoran Liao, Pengyue Wang, Shuoyu Chen et al.

2026-08-03 43
cs.RO 2608.01066

OC-VLA++: Monocular Geometry-Guided Cross-View Consistency for Viewpoint-Robust Robotic Manipulation

OC-VLA++ enhances viewpoint robustness in monocular robotic manipulation via geometry-guided cross-view supervision and action equivariance, achieving 8-15% success rate improvements under large camera shifts.

Tianyi Zhang, Ziyang Gong, Zhenjie Yang et al.

2026-08-02 49
cs.RO 2607.29302

BWM: A Low-Cost High-Fidelity World Simulator for Robot Learning

BWM integrates action-aligned data construction with diffusion-based autoregressive prediction, enhancing robot simulation fidelity and policy evaluation.

BWM Team

2026-07-31 38
Prev 1 2 3 4 5 6 7 8 ... 47 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home