cs.CL 2303.18223

A Survey of Large Language Models

Survey of large language models revealing their powerful capabilities in NLP tasks.

Wayne Xin Zhao, Kun Zhou, Junyi Li et al.

2023-04-01 38
cs.CL 2303.11315

Context-faithful Prompting for Large Language Models

Enhance LLMs' contextual faithfulness using opinion-based prompts and counterfactual demonstrations, significantly reducing memorization ratio.

Wenxuan Zhou, Sheng Zhang, Hoifung Poon et al.

2023-03-21 3
cs.CL 2303.13375

Capabilities of GPT-4 on Medical Challenge Problems

GPT-4 surpasses USMLE passing scores by over 20 points without domain-specific tuning, demonstrating strong reasoning and calibration.

Harsha Nori, Nicholas King, Scott Mayer McKinney et al.

2023-03-21 46
cs.CL 2303.08774

GPT-4 Technical Report

GPT-4 is a multimodal Transformer trained with predictive and RLHF methods, achieving human-level performance and scalable predictability.

OpenAI, Josh Achiam, Steven Adler et al.

2023-03-16 31
cs.CL 2302.13007

AugGPT: Leveraging ChatGPT for Text Data Augmentation

AugGPT leverages ChatGPT for data augmentation, significantly boosting few-shot text classification accuracy by generating diverse, semantically consistent samples.

Haixing Dai, Zhengliang Liu, Wenxiong Liao et al.

2023-02-25 51
cs.CL 2302.06100

Can GPT-3 Perform Statutory Reasoning?

Evaluated GPT-3 (text-davinci-003) on the SARA dataset for legal reasoning, using various prompting methods, achieving better results but with notable errors and knowledge gaps.

Andrew Blair-Stanek, Nils Holzenberger, Benjamin Van Durme

2023-02-13 27