1 paper · 1 filter
Michael Shalyt, Rotem Elimelech, Ido Kaminer
Large language models (LLMs) are increasingly applied to symbolic mathematics, yet existing evaluations often conflate pattern memorization with genuine reasoning. To address this…