1 paper · 1 filter
Ik-hwan Kim, Hyeongrok Han, Mingi Jung +5
Large Language Models (LLMs) often produce incorrect answers on multi-hop question answering even when the reasoning trace already contains a correct intermediate conclusion. We at…