1 paper · 1 filter
Nikolaus Holzer, William Fishell, Baishakhi Ray +1
Current training paradigms, optimized for long-horizon reasoning trace execution, have made Large Language Models (LLMs) excel at pattern matching and forward simulation of reasoni…