cs.LG 1803.02965

A Multi-Objective Deep Reinforcement Learning Framework

A DQN-based MODRL framework supports single/multi-policy and linear/nonlinear selection, recovering Pareto solutions on two benchmark tasks.

Thanh Thi Nguyen, Ngoc Duy Nguyen, Peter Vamplew et al.

2018-03-08 20
cs.CL 1803.02324

Annotation Artifacts in Natural Language Inference Data

Using fastText model to analyze annotation artifacts in NLI datasets, finding 67% of SNLI and 53% of MultiNLI data can be classified by hypothesis alone.

Suchin Gururangan, Swabha Swayamdipta, Omer Levy et al.

2018-03-07 48
cs.CL 1803.02155

Self-Attention with Relative Position Representations

Introduces relative position representations into self-attention, improving machine translation BLEU scores by 1.3 and 0.3, replacing absolute encodings.

Peter Shaw, Jakob Uszkoreit, Ashish Vaswani

2018-03-06 54
quant-ph 1803.00745

Quantum Circuit Learning

Quantum Circuit Learning combines classical and quantum computing to approximate nonlinear functions.

Kosuke Mitarai, Makoto Negoro, Masahiro Kitagawa et al.

2018-03-02 49
stat.ML 1803.00567

Computational Optimal Transport

Efficient numerical methods for optimal transport enable scalable high-dimensional distribution matching, benefiting image processing and machine learning.

Gabriel Peyré, Marco Cuturi

2018-03-02 29
cs.AI 1802.09477

Addressing Function Approximation Error in Actor-Critic Methods

TD3 algorithm employs twin critics with minimum value selection, delayed policy updates, and target smoothing, reducing overestimation bias in continuous control tasks, outperforming DDPG.

Scott Fujimoto, Herke van Hoof, David Meger

2018-02-27 70
stat.ML 1802.07073

Robust Maximization of Non-Submodular Objectives

Introduces Oblivious-Greedy for non-submodular maximization under element removal, achieving constant-factor approximation for support selection and variance reduction.

Ilija Bogunovic, Junyao Zhao, Volkan Cevher

2018-02-20 49
stat.ML 1802.05983

Disentangling by Factorising

FactorVAE improves disentanglement over β-VAE by encouraging factorial representation distribution.

Hyunjik Kim, Andriy Mnih

2018-02-16 5