3 papers
cs.AI2026
ReMath: Benchmarking Theorem Retrieval in Research-Level Mathematics
Zicheng Lyu, Wenjie Yang, Shengzhong Zhang +1
Large language models are increasingly capable at closed-world mathematical reasoning, but research assistance also requires source-grounded use of the literature. When a proof rea…
cs.CV2025
Poivre: Self-Refining Visual Pointing with Reinforcement Learning
Wenjie Yang, Zengfeng Huang
Visual pointing, which aims to localize a target by predicting its coordinates on an image, has emerged as an important problem in the realm of vision-language models (VLMs). Despi…
cs.CL2025
Right Is Not Enough: The Pitfalls of Outcome Supervision in Training LLMs for Math Reasoning
Jiaxing Guo, Wenjie Yang, Shengzhong Zhang +4
Outcome-rewarded Large Language Models (LLMs) have demonstrated remarkable success in mathematical problem-solving. However, this success often masks a critical issue: models frequ…