1 paper · 1 filter
Tong Wu, Chong Xiang, Jiachen T. Wang +4
Recently, Zaremba et al. demonstrated that increasing inference-time computation improves robustness in large proprietary reasoning LLMs. In this paper, we first show that smaller-…