1 paper · 1 filter
Giannis Chatziveroglou, Richard Yun, Maura Kelleher
This study investigates the reasoning robustness of large language models (LLMs) on mathematical problem-solving tasks under systematically introduced input perturbations. Using th…