2 papers
cs.CL2025
Right Is Not Enough: The Pitfalls of Outcome Supervision in Training LLMs for Math Reasoning
Jiaxing Guo, Wenjie Yang, Shengzhong Zhang +4
Outcome-rewarded Large Language Models (LLMs) have demonstrated remarkable success in mathematical problem-solving. However, this success often masks a critical issue: models frequ…
cs.LG2024
Your Graph Recommender is Provably a Single-view Graph Contrastive Learning
Wenjie Yang, Shengzhong Zhang, Jiaxing Guo +1
Graph recommender (GR) is a type of graph neural network (GNNs) encoder that is customized for extracting information from the user-item interaction graph. Due to its strong perfor…