2 citations · 3 across the 6 of their papers we have counts for
4 papers · 1 filter
Reactivating Test-Time Scaling for Plane Geometry Problem Solving
Xiaoqiang Kang, Shengen Wu, Maizhen Ning +5
Plane geometry problem (PGP) solving has become a critical benchmark for multimodal reasoning because it requires accurate visual perception and precise multi-step symbolic deducti…
Can GRPO Boost Complex Multimodal Table Understanding?
Xiaoqiang Kang, Shengen Wu, Zimu Wang +7
Existing table understanding methods face challenges due to complex table structures and intricate logical reasoning. While supervised finetuning (SFT) dominates existing research,…
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
Zihao Zhou, Shudong Liu, Maizhen Ning +6
Exceptional mathematical reasoning ability is one of the key features that demonstrate the power of large language models (LLMs). How to comprehensively define and evaluate the mat…
MathAttack: Attacking Large Language Models Towards Math Solving Ability
Zihao Zhou, Qiufeng Wang, Mingyu Jin +6
With the boom of Large Language Models (LLMs), the research of solving Math Word Problem (MWP) has recently made great progress. However, there are few studies to examine the secur…