cs.RO 2505.03729

Visual Imitation Enables Contextual Humanoid Control

VIDEOMIMIC converts monocular videos into environment-conditioned humanoid control policies, enabling robots to perform complex behaviors like stair climbing and sitting in diverse settings.

Arthur Allshire, Hongsuk Choi, Junyi Zhang et al.

2025-05-07 37
cs.CL 2505.02387

RM-R1: Reward Modeling as Reasoning

RM-R1 formulates reward modeling as a reasoning task using chain-of-thought and Rubrics, outperforming larger models by up to 4.9%.

Xiusi Chen, Gaotang Li, Ziqi Wang et al.

2025-05-05 38
cs.NE 2505.01262

Thinking Outside the Template with Modular GP-GOMEA

Proposes Modular GP-GOMEA with multi-tree evolution and subexpression reuse, boosting symbolic regression accuracy and interpretability.

Joe Harrison, Peter A. N. Bosman, Tanja Alderliesten

2025-05-02 49
cs.RO 2505.00693

Robotic Visual Instruction

Proposes RoVI for hand-drawn visual instructions, integrated with VIEW pipeline, achieving 87.5% success in complex robotic tasks.

Yanbang Li, Ziyang Gong, Haoyang Li et al.

2025-05-02 50
cs.CL 2505.00662

DeepCritic: Deliberate Critique with Large Language Models

DeepCritic introduces a two-stage framework leveraging Qwen2.5-72B-Instruct to enhance mathematical critique, outperforming GPT-4o and DeepSeek with significant accuracy gains.

Wenkai Yang, Jingwen Chen, Yankai Lin et al.

2025-05-02 62