1 paper
Yuyang Jiang, Chacha Chen, Teng Wu +4
Debate, where AI agents argue opposing positions, has emerged as a key approach to scalable oversight. However, debate faces a fundamental tension: models are incentivized to be pe…