2 papers
cs.LG2025
Training a Generally Curious Agent
Fahim Tajwar, Yiding Jiang, Abitha Thankaraj +4
Efficient exploration is essential for intelligent systems interacting with their environment, but existing language models often fall short in scenarios that require strategic inf…
cs.LG2025
Looking beyond the next token
Abitha Thankaraj, Yiding Jiang, J. Zico Kolter +1
The structure of causal language model training assumes that each token can be accurately predicted from the previous context. This contrasts with humans' natural writing and reaso…