cs.LG 2407.04622

On scalable oversight with weak LLMs judging strong LLMs

This paper introduces debate as a scalable oversight protocol, using weak LLMs to judge strong LLMs, showing superior performance in information-asymmetry tasks with up to 8% accuracy gain.

Zachary Kenton, Noah Y. Siegel, János Kramár et al.

2024-07-06 102 citations 44