1 paper · 1 filter
Chenrui Fan, Ming Li, Lichao Sun +1
We find that the response length of reasoning LLMs, whether trained by reinforcement learning or supervised learning, drastically increases for ill-posed questions with missing pre…