ManiTwin: Scaling Data-Generation-Ready Digital Object Dataset to 100K
ManiTwin generates 100K high-quality 3D digital assets from a single image for large-scale robotic manipulation data generation.
Kaixuan Wang, Tianxing Chen, Jiawei Liu et al.
ManiTwin generates 100K high-quality 3D digital assets from a single image for large-scale robotic manipulation data generation.
Kaixuan Wang, Tianxing Chen, Jiawei Liu et al.
BrickSim is a physics-based simulator for real-time simulation of brick assemblies, achieving 100% accuracy.
Haowei Wen, Ruixuan Liu, Weiyi Piao et al.
Achieved superior drone interception using PPO-based competitive reinforcement learning with high catch rates.
Timothée Gavin, Simon Lacroix, Murat Bronz
Proposes SignNav and START model, using spatial-temporal Transformer for semantic indoor navigation, achieving 80% success rate.
Jian Sun, Yuming Huang, He Li et al.
RAISE method ensures safety for VLA driving systems, enhancing system trust.
Gerhard Yu, Fuyuki Ishikawa, Oluwafemi Odu et al.
The study explores deployment constraints and mitigation strategies for foundation models on edge devices, emphasizing memory bandwidth and compute latency.
Utkarsh Grover, Ravi Ranjan, Mingyang Mao et al.
PRIMO R1 transforms video MLLMs into active 'Critics' using reinforcement learning, achieving 67.0% accuracy on RoboFail benchmark.
Yibin Liu, Yaxing Lyu, Daqi Gao et al.
OmniClone, Transformer-based humanoid teleoperation system, reduces MPJPE by over 66%, enabling versatile, high-fidelity control with minimal data on consumer GPU.
Yixuan Li, Le Ma, Yutang Lin et al.
PaIR-Drive enhances autonomous driving via a parallel framework, achieving 91.2 PDMS.
Zhexi Lian, Haoran Wang, Xuerun Yan et al.
Proposes GraspADMM, using ADMM to optimize multi-objective dexterous grasping, achieving 15% success rate improvement.
Liangwang Ruan, Jiayi Chen, He Wang et al.
This study reveals that a few attention heads within a frozen VLA model can detect path deviations in real time without additional training, enhancing robot navigation safety.
Jaehwan Jeong, Evelyn Zhu, Jinying Lin et al.
Proposed a feasibility-enhanced control barrier function method, significantly reducing infeasibility and improving collision avoidance in multi-UAV scenarios.
Qishen Zhong, Junlong Wu, Jian Yang et al.
Qwen2.5-VL excels in spatial reasoning for robot motion with 71.4% zero-shot accuracy.
Wenxi Wu, Jingjing Zhang, Martim Brandão
SldprtNet is a large-scale multimodal dataset with 242,000 industrial parts for semantic-driven CAD modeling.
Ruogu Li, Sikai Li, Yao Mu et al.
Significantly reduces manipulator end-effector deviation under cyberattacks using a novel active defense method.
Gabriele Gualandi, Alessandro V. Papadopoulos
Through resource-centric route fragmentation, improve multi-robot path planning efficiency in agricultural environments, achieving 95% task throughput.
James R. Heselden, Gautham P. Das
RoboStream employs Spatio-Temporal Fusion Tokens and Causal Spatio-Temporal Graphs to enable persistent memory and reasoning, achieving 90.5% success in long-horizon robotic tasks.
Yuzhi Huang, Jie Wu, Weijue Bu et al.
MotionAnymesh transforms static 3D meshes into simulation-ready digital twins using physics-constrained joint estimation and motion optimization.
WenBo Xu, Liu Liu, Li Zhang et al.
$Ψ_0$ model achieves 40% performance improvement using only 800 hours of human video and 30 hours of robot data.
Songlin Wei, Hongyi Jing, Boqian Li et al.
HumDex system uses IMU tracking and learning methods for portable humanoid dexterous manipulation, enhancing data collection efficiency and generalization.
Liang Heng, Yihe Tang, Jiajun Xu et al.