cs.AI 1910.03137

Detecting AI Trojans Using Meta Neural Analysis

Proposes Meta Neural Trojan Detection (MNTD) using jumbo learning, achieving 97% AUC in black-box Trojan detection across diverse datasets.

Xiaojun Xu, Qi Wang, Huichen Li et al.

2019-10-08 63
cs.AI 1909.10838

Talk2Car: Taking Control of Your Self-Driving Car

Introduced Talk2Car dataset with 11,959 natural language commands for urban scene target recognition; evaluated state-of-the-art models achieving up to 50.51% IoU.

Thierry Deruyttere, Simon Vandenhende, Dusan Grujicic et al.

2019-09-24 66
cs.AI 1909.05398

Interactive Fiction Games: A Colossal Adventure

Introduced Jericho environment to study language agents in interactive fiction games, achieving a 10.7% score improvement.

Matthew Hausknecht, Prithviraj Ammanabrolu, Marc-Alexandre Côté et al.

2019-09-12 4
cs.AI 1907.08194

Neural Probabilistic Logic Programming in DeepProbLog

DeepProbLog integrates neural networks with probabilistic logic programming, enabling symbolic and subsymbolic inference, program induction, and end-to-end training, advancing neuro-symbolic AI.

Robin Manhaeve, Sebastijan Dumančić, Angelika Kimmig et al.

2019-07-18 795 citations 45
cs.AI 1902.09725

Conservative Agency via Attainable Utility Preservation

Proposes Attainable Utility Preservation (AUP) to reduce reward misspecification risks by preserving multi-objective optimization capabilities.

Alexander Matt Turner, Dylan Hadfield-Menell, Prasad Tadepalli

2019-02-26 44
cs.AI 1807.06757

On Evaluation of Embodied Navigation Agents

The paper introduces SPL metric to standardize evaluation of navigation tasks, enhancing research on agents in 3D environments.

Peter Anderson, Angel Chang, Devendra Singh Chaplot et al.

2018-07-18 3
cs.AI 1802.09477

Addressing Function Approximation Error in Actor-Critic Methods

TD3 algorithm employs twin critics with minimum value selection, delayed policy updates, and target smoothing, reducing overestimation bias in continuous control tasks, outperforming DDPG.

Scott Fujimoto, Herke van Hoof, David Meger

2018-02-27 70
cs.AI 1801.00690

DeepMind Control Suite

DeepMind Control Suite offers standardized continuous control benchmarks with MuJoCo, supporting multiple algorithms and task types for RL research.

Yuval Tassa, Yotam Doron, Alistair Muldal et al.

2018-01-02 83
cs.AI 1711.09048

A Compression-Inspired Framework for Macro Discovery

Proposes a compression-based macro discovery framework that extracts, evaluates, and diversifies action sequences to accelerate RL learning in related tasks.

Francisco M. Garcia, Bruno C. da Silva, Philip S. Thomas

2017-11-25 38
cs.AI 1710.03740

Mixed Precision Training

Proposes mixed precision training using FP16 storage, FP32 master weights, and loss scaling, achieving accuracy parity with FP32 while halving memory usage.

Paulius Micikevicius, Sharan Narang, Jonah Alben et al.

2017-10-11 32