cs.CL 2212.10071

Large Language Models Are Reasoning Teachers

Proposes Fine-tune-CoT, using large models to generate reasoning samples, greatly enhancing small model reasoning performance.

Namgyu Ho, Laura Schmid, Se-Young Yun

2022-12-20 40
cs.CL 2212.09746

Evaluating Human-Language Model Interaction

HALIE framework evaluates human-AI interaction, revealing that superior non-interactive metrics do not guarantee better user experience.

Mina Lee, Megha Srivastava, Amelia Hardy et al.

2022-12-20 44
cs.LG 2212.09720

The case for 4-bit precision: k-bit Inference Scaling Laws

This study establishes that 4-bit quantization offers near-universal optimality for zero-shot performance and model size trade-offs in large language models, validated by 35,000 experiments.

Tim Dettmers, Luke Zettlemoyer

2022-12-20 40
cs.CL 2212.09611

Optimizing Prompts for Text-to-Image Generation

Proposed PROMPTIST, a reinforcement learning framework, improves text-to-image prompts, boosting reward by over 30% on Stable Diffusion.

Yaru Hao, Zewen Chi, Li Dong et al.

2022-12-20 50
cs.RO 2212.08148

Collision Avoidance Testing of the Waymo Automated Driving System

Waymo's Collision Avoidance Testing (CAT) framework evaluates SAE Level 4 ADS safety in urgent scenarios via virtual simulation and behavior models, showing a 15% collision reduction.

Kristofer D. Kusano, Kurt Beatty, Scott Schnelle et al.

2022-12-16 26 citations 40