3 citations · 5 across the 11 of their papers we have counts for
1 paper · 1 filter
Zijun Gao, Zhikun Xu, Xiao Ye +1
Large language models (LLMs) often solve challenging math exercises yet fail to apply the concept right when the problem requires genuine understanding. Popular Reinforcement Learn…