cs.CV 2104.14222

Privacy-Preserving Portrait Matting

Introduced P3M-10k dataset and P3M-Net for privacy-preserving portrait matting, achieving state-of-the-art results with face-blurred images.

Jizhizi Li, Sihan Ma, Jing Zhang et al.

2021-04-29 31
cs.CV 2104.12756

InfographicVQA

Introduces InfographicVQA dataset and evaluates Transformer-based models, revealing significant performance gaps in understanding complex infographics.

Minesh Mathew, Viraj Bagal, Rubèn Pérez Tito et al.

2021-04-27 53
cs.CV 2104.00743

Towards General Purpose Vision Systems

GPV-1 unifies vision tasks through language-conditioned shared heads, reaching 62.5 VQA and 1.023 CIDEr-D on COCO.

Tanmay Gupta, Amita Kamath, Aniruddha Kembhavi et al.

2021-04-02 17