3 citations · 5 across the 11 of their papers we have counts for
1 paper · 1 filter
Zeyu Liu, Yuhang Liu, Guanghao Zhu +9
Recent advancements in large language models (LLMs) have demonstrated substantial progress in reasoning capabilities, such as DeepSeek-R1, which leverages rule-based reinforcement…