cs.CV 2204.01955

Autoregressive 3D Shape Generation via Canonical Mapping

Proposed an autoregressive 3D shape generation method using Transformers, achieving efficient generation via semantically aligned shape composition sequences.

An-Chieh Cheng, Xueting Li, Sifei Liu et al.

2022-04-05 4
cs.CV 2203.12119

Visual Prompt Tuning

This paper introduces Visual Prompt Tuning (VPT), which inserts less than 1% trainable prompts into the input space of frozen vision transformers, outperforming full fine-tuning on 20 out of 24 tasks with significantly fewer parameters.

Menglin Jia, Luming Tang, Bor-Chun Chen et al.

2022-03-23 2841 citations 36
cs.CV 2203.09692

Facial Geometric Detail Recovery via Implicit Representation

This paper introduces a single-image facial detail recovery method using implicit surface representation and StyleGAN-based texture completion, achieving state-of-the-art results.

Xingyu Ren, Alexandros Lattas, Baris Gecer et al.

2022-03-18 55 citations 37
cs.CV 2203.09517

TensoRF: Tensorial Radiance Fields

TensoRF employs tensor decomposition of 4D feature tensors for fast, compact scene radiance field reconstruction, outperforming NeRF in speed and size.

Anpei Chen, Zexiang Xu, Andreas Geiger et al.

2022-03-18 54
cs.CV 2203.06173

Masked Visual Pre-training for Motor Control

Self-supervised MAE pretraining on large-scale real images improves pixel-based robot control, outperforming supervised encoders by up to 80%.

Tete Xiao, Ilija Radosavovic, Trevor Darrell et al.

2022-03-12 33
cs.CV 2202.08526

Point Cloud Generation with Continuous Conditioning

Proposes a continuous conditional GAN (CC-GAN) for 3D point cloud generation, achieving explicit size control with low regression error (0.28%) and superior quality (FPD 1.5290).

Larissa T. Triess, Andre Bühler, David Peter et al.

2022-02-17 43