1 paper
Yueqing Hu, Xinyang Peng, Shuting Peng +2
Recent Large Reasoning Models trained via reinforcement learning exhibit a "natural" alignment with human cognitive costs. However, we show that the prevailing paradigm of reasonin…