1 paper
Elizabeth Pavlova, Mariia Koroliuk, Karthik Viswanathan +3
We propose a new architectural change, and post-training pipeline, for making LLMs more verbose reasoners by teaching a model to truncate forward passes early. We augment an existi…