1 paper
Awni Altabaa, John Lafferty
A language model pI^¸(y∣x) trained on reasoning tasks learns to solve problems via multiple distinct strategies, yet these strategies are implicit and entangled within the m…