cs.CV 2411.00639

Event-guided Low-light Video Semantic Segmentation

EVSNet leverages event-based motion features with a lightweight fusion framework, achieving 34.1% mIoU on low-light VSPW, 11× parameter efficiency over SOTA.

Zhen Yao, Mooi Choo Chuah

2024-11-01 23 citations 45
cs.RO 2410.24221

EgoMimic: Scaling Imitation Learning via Egocentric Video

EgoMimic leverages egocentric human videos and 3D hand tracking with cross-domain alignment and joint training, boosting manipulation task success by 34-228%.

Simar Kareer, Dhruv Patel, Ryan Punamiya et al.

2024-11-01 222 citations 49