cs.CV 2205.09904

Deep transfer learning for image classification: a survey

Proposes a deep transfer learning framework with a new taxonomy, analyzing source-target data relationships to improve image classification, especially in small-data scenarios.

Jo Plested, Musa Phiri, Tom Gedeon

2022-05-20 50
cs.CV 2205.03776

SparseTT: Visual Tracking with Sparse Transformers

SparseTT employs sparse Transformer attention to improve visual tracking accuracy, achieving 40FPS with 75% faster training than TransT.

Zhihong Fu, Zehua Fu, Qingjie Liu et al.

2022-05-08 28
cs.CV 2204.08376

Detecting Deepfakes with Self-Blended Images

Self-Blended Images (SBI) enhances deepfake detection, achieving 99.64% AUC across datasets, improving cross-domain robustness.

Kaede Shiohara, Toshihiko Yamasaki

2022-04-18 42
cs.CV 2204.04676

Simple Baselines for Image Restoration

NAFNet eliminates nonlinear activations, surpasses SOTA with 33.69dB PSNR on GoPro, using only 8.4% of the original computational cost.

Liangyu Chen, Xiaojie Chu, Xiangyu Zhang et al.

2022-04-10 37
cs.CV 2204.03458

Video Diffusion Models

Video diffusion model using 3D U-Net architecture enables high-quality long video synthesis with conditional sampling and joint training.

Jonathan Ho, Tim Salimans, Alexey Gritsenko et al.

2022-04-07 75
cs.CV 2204.03444

Deep Visual Geo-localization Benchmark

Open-source benchmark framework for VG, analyzing impact of components on recall@N, model size, and efficiency.

Gabriele Berton, Riccardo Mereu, Gabriele Trivigno et al.

2022-04-07 48