stat.ML 2110.08449

Adversarial Attacks on Gaussian Process Bandits

This paper analyzes adversarial attacks on Gaussian process bandits, proposing white-box and black-box methods that successfully steer algorithms toward target regions with low budgets.

Eric Han, Jonathan Scarlett

2021-10-16 46
cs.CL 2110.08193

BBQ: A Hand-Built Bias Benchmark for Question Answering

Introduces BBQ, a handcrafted bias benchmark for question answering, revealing models' reliance on stereotypes especially under low-information conditions, with bias scores up to 100%.

Alicia Parrish, Angelica Chen, Nikita Nangia et al.

2021-10-16 51
cs.CL 2110.06341

Learning Compact Metrics for MT

Proposes RemBERT, a distilled multilingual evaluation metric reaching 92.6% of the large model's performance with only one-third parameters.

Amy Pu, Hyung Won Chung, Ankur P. Parikh et al.

2021-10-13 49
cs.CL 2110.03215

Towards Continual Knowledge Learning of Language Models

Proposes CKL framework with INVARIANTLAMA, UPDATEDLAMA, NEWLAMA datasets and FUAR metric; uses parameter expansion methods to improve knowledge retention and acquisition.

Joel Jang, Seonghyeon Ye, Sohee Yang et al.

2021-10-07 30