1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025
Unveiling Confirmation Bias in Chain-of-Thought Reasoning
Yue Wan, Xiaowei Jia, Xiang Lorraine Li
Chain-of-thought (CoT) prompting has been widely adopted to enhance the reasoning capabilities of large language models (LLMs). However, the effectiveness of CoT reasoning is incon…
cs.CL2024★ 1 cited
Every Answer Matters: Evaluating Commonsense with Probabilistic Measures
Qi Cheng, Michael Boratko, Pranay Kumar Yelugam +4
Large language models have demonstrated impressive performance on commonsense tasks; however, these tasks are often posed as multiple-choice questions, allowing models to exploit s…