1 paper
Tej Deep Pala, Panshul Sharma, Amir Zadeh +2
Large Language Models (LLMs) are prone to hallucination, especially during multi-hop and reasoning-intensive tasks such as mathematical problem solving. While Outcome Reward Models…