2 papers
cs.AI2026
MedCalc-R1: Knowledge-Guided Reward Framework for Medical Mathematical Reasoning
Haotian Wang, Lian Yan, Xingzhi Yao +4
In Reinforcement Learning with Verifiable Rewards (RLVR) frameworks for mathematical reasoning tasks, floating-point results are typically evaluated using a tolerance-based reward.…
cs.CL2025
AgriEval: A Comprehensive Chinese Agricultural Benchmark for Large Language Models
Lian Yan, Haotian Wang, Chen Tang +5
In the agricultural domain, the deployment of large language models (LLMs) is hindered by the lack of training data and evaluation benchmarks. To mitigate this issue, we propose Ag…