Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Learning to Reason in 13 Parameters
John X. Morris, Niloofar Mireshghallah, Mark Ibrahim +1
Recent research has shown that language models can learn to \textit{reason}, often via reinforcement learning. Some work even trains low-rank parameterizations for reasoning, but c…
cs.LG2024
Do language models plan ahead for future tokens?
Wilson Wu, John X. Morris, Lionel Levine
Do transformers "think ahead" during inference at a given position? It is known transformers prepare information in the hidden states of the forward pass at time step that is t…