1 paper
Charlie Snell, Jaehoon Lee, Kelvin Xu +1
Enabling LLMs to improve their outputs by using more test-time computation is a critical step towards building generally self-improving agents that can operate on open-ended natura…