cs.CV 1504.06375

Holistically-Nested Edge Detection

Introduced HED, a new edge detection algorithm achieving an ODS F-score of 0.782 on the BSD500 dataset.

Saining Xie, Zhuowen Tu

2015-04-24 46
stat.ML 1504.02338

Kernel Manifold Alignment

Kernel Manifold Alignment (KEMA) enables multi-source, unpaired domain alignment with superior performance on synthetic and real datasets.

Devis Tuia, Gustau Camps-Valls

2015-04-09 50
cs.LG 1504.00702

End-to-End Training of Deep Visuomotor Policies

Proposes end-to-end training of deep visuomotor policies using Guided Policy Search with a 92,000-parameter CNN for direct image-to-torque mapping.

Sergey Levine, Chelsea Finn, Trevor Darrell et al.

2015-04-03 33
cs.LG 1503.03578

LINE: Large-scale Information Network Embedding

LINE efficiently embeds large-scale networks by optimizing first- and second-order proximities with edge sampling, handling millions of nodes and billions of edges.

Jian Tang, Meng Qu, Mingzhe Wang et al.

2015-03-12 56
cs.CV 1503.03167

Deep Convolutional Inverse Graphics Network

DC-IGN learns interpretable image representations using SGVB, generating images with varied poses and lighting.

Tejas D. Kulkarni, Will Whitney, Pushmeet Kohli et al.

2015-03-11 7
stat.ML 1503.02531

Distilling the Knowledge in a Neural Network

Knowledge distillation transfers ensemble model knowledge into a single small model, achieving near-ensemble performance on MNIST and speech recognition tasks.

Geoffrey Hinton, Oriol Vinyals, Jeff Dean

2015-03-09 25922 citations 54
cs.LG 1502.07073

Strongly Adaptive Online Learning

Proposes a meta-algorithm (SAOL) transforming low-regret algorithms into strongly adaptive ones, ensuring near-optimal performance on every interval with \( O(\log T) \) overhead.

Amit Daniely, Alon Gonen, Shai Shalev-Shwartz

2015-02-25 39
cs.LG 1502.05477

Trust Region Policy Optimization

TRPO (Trust Region Policy Optimization) guarantees monotonic policy improvement using KL constraints, excelling in large neural network policy training for robotics and Atari games.

John Schulman, Sergey Levine, Philipp Moritz et al.

2015-02-19 8275 citations 49
cs.LG 1502.04681

Unsupervised Learning of Video Representations using LSTMs

Proposes a multi-layer LSTM encoder-decoder framework for unsupervised video representation learning, improving action recognition especially with limited labeled data.

Nitish Srivastava, Elman Mansimov, Ruslan Salakhutdinov

2015-02-17 58