cs.CL 2006.15020

Pre-training via Paraphrasing

MARGE introduces a retrieval-based pretraining method for multilingual multi-document paraphrasing, achieving strong zero-shot performance on translation and summarization.

Mike Lewis, Marjan Ghazvininejad, Gargi Ghosh et al.

2020-06-26 60
cs.CL 2006.14799

Evaluation of Text Generation: A Survey

This survey reviews three categories of NLG evaluation metrics: human, automatic, and machine-learned, analyzing their strengths, weaknesses, and future directions.

Asli Celikyilmaz, Elizabeth Clark, Jianfeng Gao

2020-06-26 63
cs.CV 2006.11275

Center-based 3D Object Detection and Tracking

CenterPoint uses point-based detection with heatmaps and regression, achieving 65.5 NDS on nuScenes with a single model.

Tianwei Yin, Xingyi Zhou, Philipp Krähenbühl

2020-06-20 21
cs.LG 2006.10598

Neural Parameter Allocation Search

Proposes NPAS and SSNs for automatic parameter-efficient neural network training under arbitrary budgets.

Bryan A. Plummer, Nikoli Dryden, Julius Frost et al.

2020-06-18 42
stat.ML 2006.10459

Stochastic bandits with arm-dependent delays

PatientBandits algorithm excels in handling stochastic bandits with arm-dependent delays.

Anne Gael Manegueu, Claire Vernade, Alexandra Carpentier et al.

2020-06-18 5
cs.CV 2006.10204

BlazePose: On-device Real-time Body Pose tracking

BlazePose is a lightweight CNN for real-time human pose estimation on mobile devices, outputting 33 keypoints at over 30fps.

Valentin Bazarevsky, Ivan Grishchenko, Karthik Raveendran et al.

2020-06-18 47
cs.LG 2006.10029

Big Self-Supervised Models are Strong Semi-Supervised Learners

This paper introduces semi-supervised learning with Big Self-Supervised Models (SimCLRv2), achieving 73.9% top-1 accuracy on ImageNet with only 1% labels, outperforming previous methods by 10x.

Ting Chen, Simon Kornblith, Kevin Swersky et al.

2020-06-18 2601 citations 43
cs.LG 2006.09503

Memory-Efficient Pipeline-Parallel DNN Training

Proposes PipeDream-2BW, a memory-efficient pipeline training system for large-scale DNNs, achieving up to 20× speedup via double-buffered weight updates and automated model partitioning.

Deepak Narayanan, Amar Phanishayee, Kaiyu Shi et al.

2020-06-17 307 citations 46