cs.CV 2503.21776

Video-R1: Reinforcing Video Reasoning in MLLMs

Introduces Video-R1 with T-GRPO algorithm and hybrid datasets, achieving 37.1% accuracy on VSI-Bench, surpassing GPT-4o.

Kaituo Feng, Kaixiong Gong, Bohao Li et al.

2025-03-28 35