1 paper
Quoc Tuan Pham, Mehdi Jafari, Flora Salim
Test-time compute is central to large reasoning models, yet analysing their reasoning behaviour through generated text is increasingly impractical and unreliable. Response length i…