1 paper
Riccardo Revalor, Jalees Rehman, Debjit Pal
Large-Language Models (LLMs) can be prone to flawed and unfaithful reasoning that decoding strategies like Self-Consistency (SC) fail to detect as they evaluate only final-answer a…