1 paper
Wei Shen, Han Wang, Haoyu Li +1
Large Language Models (LLMs) have been demonstrating strong reasoning capability with their chain-of-thoughts (CoT), which are routinely used by humans to judge answer quality. Thi…