cs.IR 2502.07303

Flow Matching for Collaborative Filtering

FlowCF enhances collaborative filtering accuracy using flow matching, achieving fastest inference speeds in experiments.

Chengkai Liu, Yangtian Zhang, Jianling Wang et al.

2025-02-11 19
cs.CV 2502.04896

Goku: Flow Based Video Generative Foundation Models

Goku employs rectified flow Transformer for joint image-video generation, achieving top-tier performance with 0.76 on GenEval and 84.85 on VBench.

Shoufa Chen, Chongjian Ge, Yuqi Zhang et al.

2025-02-07 46
cs.LG 2502.04809

Humans Coexist, So Must Embodied Artificial Agents

Proposes the concept of coexistence for embodied agents, emphasizing continuous adaptation leveraging situated knowledge for long-term human interaction.

Hannah Kuehn, Joseph La Delfa, Miguel Vasco et al.

2025-02-07 49
cs.RO 2502.04584

Joint State and Noise Covariance Estimation

Proposes a convex-structured joint estimation method for states and noise covariance, with analytical solutions, applied to SLAM and robotics.

Kasra Khosoussi, Iman Shames

2025-02-07 45
cs.CV 2502.04507

Fast Video Generation with Sliding Tile Attention

Introduces Sliding Tile Attention (STA), achieving 2.8-17× speedup in video diffusion models with minimal quality loss, based on local 3D attention patterns.

Peiyuan Zhang, Yongqi Chen, Runlong Su et al.

2025-02-07 28
cs.LG 2502.04468

Iterative Importance Fine-tuning of Diffusion Models

Proposes iterative importance fine-tuning of diffusion models via self-supervision, optimizing control for conditional sampling with theoretical guarantees.

Alexander Denker, Shreyas Padhy, Francisco Vargas et al.

2025-02-07 55
cs.LG 2502.03461

Do Large Language Model Benchmarks Test Reliability?

Introduces ‘Platinum Benchmarks’ to reduce label noise, assesses LLM reliability, revealing even state-of-the-art models fail on simple tasks with 5% error rate.

Joshua Vendrow, Edward Vendrow, Sara Beery et al.

2025-02-06 53 citations 49
cs.CL 2502.03387

LIMO: Less is More for Reasoning

LIMO achieves complex reasoning with minimal data, scoring 63.3% on AIME24 and 95.6% on MATH500.

Yixin Ye, Zhen Huang, Yang Xiao et al.

2025-02-06 43