3 citations · 3 across the 1 of their papers we have counts for
1 paper
Zhenting Qi, Mingyuan Ma, Jiahang Xu +3
This paper introduces rStar, a self-play mutual reasoning approach that significantly improves reasoning capabilities of small language models (SLMs) without fine-tuning or superio…