2 papers
cs.LG2026
SenTSR-Bench: Thinking with Injected Knowledge for Time-Series Reasoning
Zelin He, Boran Han, Xiyuan Zhang +10
Time-series diagnostic reasoning is essential for many applications, yet existing solutions face a persistent gap: general reasoning large language models (GRLMs) possess strong re…
cs.LG2025
From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning
Yuzhen Huang, Weihao Zeng, Xingshan Zeng +2
Trustworthy verifiers are essential for the success of reinforcement learning with verifiable reward (RLVR), which is the core methodology behind various large reasoning models suc…