cs.CV 2204.04676

Simple Baselines for Image Restoration

NAFNet eliminates nonlinear activations, surpasses SOTA with 33.69dB PSNR on GoPro, using only 8.4% of the original computational cost.

Liangyu Chen, Xiaojie Chu, Xiangyu Zhang et al.

2022-04-10 39
cs.CV 2204.03458

Video Diffusion Models

Video diffusion model using 3D U-Net architecture enables high-quality long video synthesis with conditional sampling and joint training.

Jonathan Ho, Tim Salimans, Alexey Gritsenko et al.

2022-04-07 79
cs.CV 2204.03444

Deep Visual Geo-localization Benchmark

Open-source benchmark framework for VG, analyzing impact of components on recall@N, model size, and efficiency.

Gabriele Berton, Riccardo Mereu, Gabriele Trivigno et al.

2022-04-07 52
cs.CL 2204.02311

PaLM: Scaling Language Modeling with Pathways

Pathways系统支持下的540B参数PaLM模型,显著提升少样学习能力,超越多项自然语言任务的SOTA,展现出大规模模型的潜力。

Aakanksha Chowdhery, Sharan Narang, Jacob Devlin et al.

2022-04-06 8306 citations 54
cs.CV 2204.01955

Autoregressive 3D Shape Generation via Canonical Mapping

Proposed an autoregressive 3D shape generation method using Transformers, achieving efficient generation via semantically aligned shape composition sequences.

An-Chieh Cheng, Xueting Li, Sifei Liu et al.

2022-04-05 16
cs.LG 2204.01855

A Survey on Graph Representation Learning Methods

This survey systematically reviews graph embedding methods, including traditional and GNN-based techniques for static and dynamic graphs, analyzing over 300 papers since 2017.

Shima Khoshraftar, Aijun An

2022-04-05 266 citations 31