cs.CL 1907.05242

Large Memory Layers with Product Keys

Introduces Product Keys-based structured memory layer, enabling billion-parameter capacity with efficient exact search, outperforming deeper transformers in language modeling.

Guillaume Lample, Alexandre Sablayrolles, Marc'Aurelio Ranzato et al.

2019-07-10 46
stat.ML 1906.11471

Deep Active Learning with Adaptive Acquisition

Introduced a deep active learning method with adaptive acquisition, showing superior performance across datasets.

Manuel Haussmann, Fred A. Hamprecht, Melih Kandemir

2019-06-27 5
cs.LG 1906.10827

Hierarchical Optimal Transport for Document Representation

Hierarchical Optimal Transport (HOTT) combines topic models and word embeddings to efficiently measure document similarity, outperforming WMD in speed with comparable accuracy.

Mikhail Yurochkin, Sebastian Claici, Edward Chien et al.

2019-06-26 65