1 paper · 1 filter
Dylan Jayabahu, Tinuade Adeleke
Reasoning models do not stop when they know the answer. On DeepSeek-R1-Distill-Qwen-7B the chain of thought runs about twice as long as the model's own answer probability takes to…