cs.CL 2202.06417

A Contrastive Framework for Neural Text Generation

Contrastive training (SimCTG) and contrastive search improve diversity and coherence in neural text generation, outperforming SOTA methods.

Yixuan Su, Tian Lan, Yan Wang et al.

2022-02-14 58
cs.CL 2202.05262

Locating and Editing Factual Associations in GPT

Proposes causal intervention and ROME method to locate and edit factual associations in GPT, achieving high success rates and better control over knowledge editing.

Kevin Meng, David Bau, Alex Andonian et al.

2022-02-11 31
cs.PL 2203.07814

Competition-Level Code Generation with AlphaCode

AlphaCode leverages large-scale Transformer models with filtering and clustering to achieve top54.3% in competitive programming, solving 34.2% of problems.

Yujia Li, David Choi, Junyoung Chung et al.

2022-02-09 44
cs.RO 2202.03631

Robotic Grasping from Classical to Modern: A Survey

Integrating classical and data-driven methods, this work proposes a multi-layered robotic grasping framework, significantly improving stability and adaptability.

Hanbo Zhang, Jian Tang, Shiguang Sun et al.

2022-02-08 65
cs.LG 2202.03528

TACTiS: Transformer-Attentional Copulas for Time Series

Proposes TACTiS, a Transformer-based model using attention copulas for joint probabilistic forecasting of high-dimensional multivariate time series.

Alexandre Drouin, Étienne Marcotte, Nicolas Chapados

2022-02-08 37
cs.LG 2202.03376

Message Passing Neural PDE Solvers

MP-PDE unifies classical local solvers with message passing and improves autoregressive stability via pushforward training.

Johannes Brandstetter, Daniel Worrall, Max Welling

2022-02-08 25
cs.CL 2202.03286

Red Teaming Language Models with Language Models

Proposes automated red teaming using language models to generate and detect harmful outputs in 280B parameter chatbots, employing multi-strategy approaches.

Ethan Perez, Saffron Huang, Francis Song et al.

2022-02-07 22