5 citations · 5 across the 4 of their papers we have counts for
1 paper · 1 filter
Qifan Zhang, Jianhao Ruan, Aochuan Chen +4
Large Reasoning Models (LRMs) have advanced rapidly; however, existing benchmarks in mathematics, code, and common-sense reasoning remain limited. They lack long-context evaluation…