cs.CL 2105.09680

KLUE: Korean Language Understanding Evaluation

KLUE constructs 8 Korean NLU tasks, using from-scratch data collection and pretrained models KLUE-BERT and KLUE-RoBERTa, outperforming multilingual and open-source baselines.

Sungjoon Park, Jihyung Moon, Sungdong Kim et al.

2021-05-20 68
cs.CL 2104.08696

Knowledge Neurons in Pretrained Transformers

Introduces the concept of knowledge neurons using integrated gradients, identifying neurons responsible for factual knowledge in BERT, enabling knowledge editing without fine-tuning, with an average of 4.13 neurons per fact.

Damai Dai, Li Dong, Yaru Hao et al.

2021-04-18 751 citations 28
cs.CL 2104.07567

Retrieval Augmentation Reduces Hallucination in Conversation

Retrieval-augmented architecture reduces hallucination by over 60%, outperforming SOTA on two knowledge-grounded tasks, enhancing factual accuracy in dialogue.

Kurt Shuster, Spencer Poff, Moya Chen et al.

2021-04-16 1199 citations 29
cs.CL 2104.07423

The Role of Context in Detecting Previously Fact-Checked Claims

Integrating source and target context modeling, including coreference resolution and multi-hop reasoning via Transformer-XH, improves detection of previously fact-checked claims by over 10 MAP points.

Shaden Shaar, Firoj Alam, Giovanni Da San Martino et al.

2021-04-15 43
cs.CL 2104.06598

AR-LSAT: Investigating Analytical Reasoning of Text

AR-LSAT combines Transformer and symbolic reasoning, revealing deep reasoning gaps in current models with accuracy below 35%. ARM achieves 65%, still behind humans.

Wanjun Zhong, Siyuan Wang, Duyu Tang et al.

2021-04-14 68 citations 34