1 paper
Iddo Drori, Gaston Longhitano, Mao Mao +11
Reasoning LLMs such as OpenAI o1, o3 and DeepSeek R1 have made significant progress in mathematics and coding, yet find challenging advanced tasks such as International Mathematica…