cs.CL 2312.06550

LLM360: Towards Fully Transparent Open-Source LLMs

Proposes LLM360 framework, fully open-sourcing 7B models, training code, data, checkpoints, and analysis for transparency.

Zhengzhong Liu, Aurick Qiao, Willie Neiswanger et al.

2023-12-12 122 citations 45
cs.CL 2312.03718

Large Language Models in Law: A Survey

First survey of legal LLMs: Transformer/ChatGPT promise efficiency, but data, ethics, and due process remain bottlenecks.

Jinqi Lai, Wensheng Gan, Jiayang Wu et al.

2023-11-26 256 citations 46
cs.CL 2311.09144

Grounding Gaps in Language Model Generations

Proposes grounding acts as metrics to evaluate dialogue, finds LLMs generate 77.5% fewer grounding acts than humans, and preference tuning reduces these acts further.

Omar Shaikh, Kristina Gligorić, Ashna Khetan et al.

2023-11-16 39