cs.CL 2401.10020

Self-Rewarding Language Models

Self-Rewarding language model using iterative DPO training, surpassing Claude 2, Gemini Pro, and GPT-4 0613 on AlpacaEval 2.0 with Llama 2 70B.

Weizhe Yuan, Richard Yuanzhe Pang, Kyunghyun Cho et al.

2024-01-18 693 citations 29
cs.CV 2401.09414

Vlogger: Make Your Dream A Vlog

Vlogger combines LLM and diffusion models to generate 5-minute coherent long videos, maintaining script-actor consistency.

Shaobin Zhuang, Kunchang Li, Xinyuan Chen et al.

2024-01-18 25
cs.CV 2401.09413

POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images

POP-3D introduces a zero-shot open-vocabulary 3D occupancy prediction model leveraging multimodal self-supervised learning, achieving 78% mIoU on nuScenes without manual annotations.

Antonin Vobecky, Oriane Siméoni, David Hurych et al.

2024-01-18 39