1 paper · 1 filter
Riccardo Revalor, Jalees Rehman, Debjit Pal
Large-Language Models (LLMs) can be prone to flawed and unfaithful reasoning that decoding strategies like Self-Consistency (SC) fail to detect as they evaluate only final-answer a…