cs.CV 2011.00931

Point Transformer

Point Transformer employs multi-head attention and SortNet to achieve permutation-invariant point cloud features, capturing local and global structures.

Nico Engel, Vasileios Belagiannis, Klaus Dietmayer

2020-11-02 27
cs.CV 2010.15614

An Overview Of 3D Object Detection

Fusion of RGB and LiDAR data using 2D detectors and pixel mapping improves 3D object detection on nuScenes, achieving ~67% mAP.

Yilin Wang, Jiayi Ye

2020-10-29 54
cs.CV 2010.10732

SCOP: Scientific Control for Reliable Neural Network Pruning

SCOP introduces knockoff features as scientific control, achieving 57.8% parameter reduction and 60.2% FLOPs reduction on ResNet-101 with only 0.01% accuracy loss.

Yehui Tang, Yunhe Wang, Yixing Xu et al.

2020-10-21 204 citations 28
cs.CV 2010.06897

Adaptive-Attentive Geolocalization from few queries: a hybrid approach

Proposes AdAGeo combining attention and few-shot unsupervised domain adaptation, achieving over 13% improvement in cross-domain visual place recognition with only 5 target images.

Gabriele Moreno Berton, Valerio Paolicelli, Carlo Masone et al.

2020-10-14 51 citations 32
cs.CV 2009.09633

3D-FUTURE: 3D Furniture shape with TextURE

Introduces 3D-FUTURE, a large-scale dataset with 20,240 synthetic indoor images and 9,992 detailed textured furniture models for multi-task 3D scene understanding.

Huan Fu, Rongfei Jia, Lin Gao et al.

2020-09-21 32
cs.CV 2008.05511

Free View Synthesis

Proposes a scene-agnostic neural view synthesis method using SfM and MVS, outperforming state-of-the-art on Tanks and Temples with over 50% LPIPS reduction.

Gernot Riegler, Vladlen Koltun

2020-08-13 42