2 papers
cs.LG2026
Decomposing Prediction Mechanisms for In-Context Recall
Sultan Daniels, Dylan Davis, Dhruv Gautam +3
We introduce a new family of toy problems that combine features of linear-regression-style continuous in-context learning (ICL) with discrete associative recall. We pretrain transf…
cs.AI2025
RefactorBench: Evaluating Stateful Reasoning in Language Agents Through Code
Dhruv Gautam, Spandan Garg, Jinu Jang +2
Recent advances in language model (LM) agents and function calling have enabled autonomous, feedback-driven systems to solve problems across various digital domains. To better unde…