1 paper
Neeraja Kirtane, Yuvraj Khanna, Peter Relan
Large language models excel on math benchmarks, but their math reasoning robustness to linguistic variation is underexplored. While recent work increasingly treats high-difficulty…