cs.CV 1708.06320

Learning Spread-out Local Feature Descriptors

Proposes a regularization method based on uniform distribution to maximize feature spread, combined with triplet loss, significantly improving local descriptor performance.

Xu Zhang, Felix X. Yu, Sanjiv Kumar et al.

2017-08-22 52
cs.CV 1708.05375

Learning a Multi-View Stereo Machine

Proposes an end-to-end learned multi-view stereo system leveraging differentiable geometric projections, achieving high-quality 3D reconstructions from few views, tested on ShapeNet.

Abhishek Kar, Christian Häne, Jitendra Malik

2017-08-18 68
cs.CV 1707.07410

Toward Geometric Deep SLAM

Proposes deep CNN-based point detector MagicPoint and homography estimator MagicWarp for robust, real-time SLAM.

Daniel DeTone, Tomasz Malisiewicz, Andrew Rabinovich

2017-07-24 62
cs.CV 1707.06484

Deep Layer Aggregation

Deep Layer Aggregation (DLA) uses iterative and hierarchical fusion to improve recognition with fewer parameters, outperforming traditional skip connections.

Fisher Yu, Dequan Wang, Evan Shelhamer et al.

2017-07-20 50
cs.CV 1707.04993

MoCoGAN: Decomposing Motion and Content for Video Generation

MoCoGAN decomposes content and motion in a latent space, enabling controllable, unsupervised video generation with improved content consistency and dynamic diversity.

Sergey Tulyakov, Ming-Yu Liu, Xiaodong Yang et al.

2017-07-17 41