1 paper · 1 filter
Tengxiao Liu, Deepak Nathani, Zekun Li +2
Recent progress in large language model (LLM) reasoning has focused on domains like mathematics and coding, where abundant high-quality data and objective evaluation metrics are re…