cs.CV 2103.16076

Face Forensics in the Wild

Constructed FFIW-10K dataset and proposed multi-instance attention model for multi-person face forgery detection.

Tianfei Zhou, Wenguan Wang, Zhiyuan Liang et al.

2021-03-30 44
cs.CV 2103.06191

A Study of Face Obfuscation in ImageNet

This study evaluates face obfuscation (blurring, overlay) on ImageNet recognition, showing less than 1% accuracy drop, supporting privacy-preserving vision.

Kaiyu Yang, Jacqueline Yau, Li Fei-Fei et al.

2021-03-11 63
cs.CV 2103.02603

Towards Open World Object Detection

Proposes ORE framework combining contrastive clustering and energy models for open world object detection, achieving state-of-the-art unknown detection and incremental learning.

K J Joseph, Salman Khan, Fahad Shahbaz Khan et al.

2021-03-04 35
cs.CV 2103.02597

Neural 3D Video Synthesis from Multi-view Video

Proposed DyNeRF with space-time latent codes achieves high-quality 10s 30FPS multi-view video synthesis, reducing training time by 40x.

Tianye Li, Mira Slavcheva, Michael Zollhoefer et al.

2021-03-04 32
cs.CV 2103.02406

Multi-attentional Deepfake Detection

Proposes a multi-attentional deepfake detection network reformulating the task as fine-grained classification, outperforming traditional binary classifiers.

Hanqing Zhao, Wenbo Zhou, Dongdong Chen et al.

2021-03-03 66
cs.CV 2102.13090

IBRNet: Learning Multi-View Image-Based Rendering

IBRNet uses multi-view feature fusion and Transformer modules to synthesize high-res novel views without scene-specific optimization, achieving state-of-the-art generalization.

Qianqian Wang, Zhicheng Wang, Kyle Genova et al.

2021-02-26 36
cs.CV 2102.12092

Zero-Shot Text-to-Image Generation

Transformers trained on 250M image-text pairs enable zero-shot text-to-image generation, outperforming previous models in diversity and realism.

Aditya Ramesh, Mikhail Pavlov, Gabriel Goh et al.

2021-02-24 66