GptGet
Features PaperForge Apps Papers Blog Contact AI Chat 中文
Sort: Latest Popular Citations
All Artificial Intelligence Computation and Language Computer Vision Information Retrieval Machine Learning Machine Learning (Stats) Neural and Evolutionary Computing Robotics
cs.RO 2607.06337

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics

Physically-grounded GPU apple-tree simulation using Euler-Bernoulli beams, rupture, and detachment mechanics, enabling autonomous harvesting research.

Humphrey Munn

2026-07-07 43
cs.RO 2607.06262

Optimal Transport Q-Learning for Flow Policy Steering and Acceleration

OTQL integrates advantage-weighted conditional optimal transport with flow models, achieving few-step inference and efficient policy fine-tuning, boosting success rates from 36% to 86%.

Andreas Sochopoulos, Esmeralda S. Whitammer, Nikolaos Tsagkas et al.

2026-07-07 48
cs.RO 2607.13059

GPUSimBench: Towards Scalable and Reliable GPU-Accelerated Simulators in Embodied AI

This paper introduces GPUSimBench, a benchmark system to evaluate GPU-based robotic simulators' scalability, physical fidelity, and non-determinism, revealing key limitations.

Huzhenyu Zhang, Shenghai Yuan, Wenrui Yan et al.

2026-07-06 44
cs.RO 2607.04880

PRISM: Personalized Robotic Dataset Generation via Image-based Scene and Motion Synthesis

PRISM generates personalized robotic datasets from a single image and instruction, constructing semantically aligned digital scenes with instance diversity.

Dogyu Ko, Haneul Kim, Chanyoung Yeo et al.

2026-07-06 51
cs.RO 2607.05468

Learning 4D Geometric Priors for Inference-Efficient World Action Models

Proposes MECo-WAM, integrating 4D geometric priors during training, boosting manipulation success to 98.2% without increasing inference cost.

Jianjun Zhang, Jian Zhu, Taiyi Su et al.

2026-07-06 49
cs.RO 2607.04714

Geometry-Aware Motion Latents for Learning Robust Manipulation Policies

GeoMoLa learns motion latents by predicting point cloud evolution, achieving SOTA performance with single-view RGB-D input.

Yunchao Zhang, Yijia Weng, Ruizhe Liu et al.

2026-07-06 0
cs.RO 2607.03941

WSA$_1$: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control

WSA1 employs 3D world-spatial-action joint modeling with only 6K hours of demonstration data, achieving 93% success in RoboTwin2.0 and +20% over SOTA in real tasks.

Jiahao Jiang, Jianing Zhang, Zhenhan Yin et al.

2026-07-05 49
cs.RO 2607.03828

ObjRetarget: An Object-Aware Motion Retargeting Framework with Anthropomorphic Arm Constraints and Polyhedral Hand Modeling

ObjRetarget integrates human videos, polyhedral hand models, and anthropomorphic constraints to achieve high-precision robot motion retargeting.

Yuanchuan Lai, Qing Gao, Ziyan Liang et al.

2026-07-04 48
cs.RO 2607.03693

CoRE-VLA: Towards Scalable and Robust Vision-Language-Action Modeling via Conditional Routing of Experts

CoRE-VLA uses conditional routing of experts for scalable and robust vision-language-action modeling, excelling in multi-task and long-horizon tasks.

Haozhe Zhang, Sixian Li, Yifei Zhang et al.

2026-07-04 5
cs.RO 2607.03570

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space

Proposed Unified Hand Action Space (UHAS) enables cross-platform dexterous manipulation; experiments show effective transfer across different hand types.

Luis Felipe Casas, Robert Teal, Keval Shah et al.

2026-07-04 36
cs.RO 2607.03163

Beyond Point-Attached Semantics: Object-Centric Semantic Fields for Generalizable Manipulation

Proposes an object-centric continuous semantic field method to enhance robot manipulation generalization.

Zheng Sun, Lerong Zhang, Zhihao Li et al.

2026-07-03 5
cs.RO 2607.02845

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation

AmpAttention boosts multi-view robotic manipulation success rate to 91%, reducing training time by 33.3%.

Jin Yang, Ping Wei, Nanning Zheng

2026-07-03 4
cs.RO 2607.02646

EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots

EVA-Client provides a unified deployment, debugging, and data collection framework for embodied policies across multiple robot platforms, integrating various real-time inference strategies.

Heqing Yang, Yang Yi, Liyao Wang et al.

2026-07-03 44
cs.RO 2607.02322

The Moving Eye: Enhancing VLA Spatial Generalization via Hybrid Dynamic Data Collection

Proposed a hybrid dynamic data collection method, significantly enhancing VLA spatial generalization with a 40% success rate increase.

Jincheng Tang, Yilong Zhu, Zhengyuan Xie et al.

2026-07-02 4
cs.RO 2607.02195

Bridge-WA: Predicting Where and How the World Changes for Robotic Action

Bridge-WA enhances robotic action success by predicting scene changes, achieving a 9.7% improvement on VLABench.

Yongjie Bai, Hanting Wang, Mingtong Dai et al.

2026-07-02 4
cs.RO 2607.01938

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation

PhysMani combines a physics-principled 3D Gaussian world model with a future-aware action policy model to enhance dynamic object manipulation success rates.

Peng Yun, Shouwang Huang, Hao Li et al.

2026-07-02 4
cs.RO 2607.01067

Human-Centric Transferable Tactile Pre-Training for Dexterous Robotic Manipulation

Introduces H-Tac dataset and TTP pre-training system, enabling cross-robot fine manipulation with 98% success rate.

Chi Zhang, Penglin Cai, Ziheng Xi et al.

2026-07-01 49
cs.RO 2607.01060

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation

RoboWorld combines STEP FORCING with autoregressive video models, achieving 0.989 correlation with real-world robot evaluation.

Byeongguk Jeon, Seonghyeon Ye, JaeHyeok Doo et al.

2026-07-01 49
cs.RO 2607.00836

From World Models to World Action Models: A Concise Tutorial for Robotics

Unified framework for world and action models integrating visual prediction and decision-making, demonstrated on robotic tasks with 20% performance boost.

Xiaoxiong Zhang, Xiong Zeng, Wei Zhang

2026-07-01 47
cs.RO 2607.00776

From Prediction Uncertainty to Conformalized Distance Fields for Safe Motion Planning

Proposes conformalized distance fields via functional prediction for safe motion planning, leveraging low-rank residuals for efficient, field-level safety guarantees.

Jaeuk Shin, Yoonseok Ra, Insoon Yang

2026-07-01 47
Prev 1 ... 4 5 6 7 8 9 10 ... 47 Next

© 2026 GptGet.net - Paper Insights Platform

Paper List Submit Paper Help GptGet Home