cs.CL 2211.09260

Task-aware Retrieval with Instructions

Proposes TART, a task-aware retrieval system using multi-task instruction tuning, excelling in zero-shot and cross-domain scenarios.

Akari Asai, Timo Schick, Patrick Lewis et al.

2022-11-17 44
cs.CL 2211.05826

The CRINGE Loss: Learning what language not to model

CRINGE loss leverages contrastive negative generation with iterative self-labeling, significantly improving safety and coherence in language models, outperforming baselines.

Leonard Adolphs, Tianyu Gao, Jing Xu et al.

2022-11-11 37
cs.CL 2210.09150

Prompting GPT-3 To Be Reliable

Simple prompts enhance GPT-3's reliability in generalizability, social bias, calibration, and factuality.

Chenglei Si, Zhe Gan, Zhengyuan Yang et al.

2022-10-17 9
cs.CL 2210.07229

Mass-Editing Memory in a Transformer

MEMIT mass-edits 10,000 facts in GPT-J, reaching an 85.8 COUNTERFACT editing score at scale.

Kevin Meng, Arnab Sen Sharma, Alex Andonian et al.

2022-10-14 19