1 citations · 1 across the 2 of their papers we have counts for
3 papers
Offline Imitation Learning from Multiple Baselines with Applications to Compiler Optimization
Teodor V. Marinov, Alekh Agarwal, Mircea Trofin
This work studies a Reinforcement Learning (RL) problem in which we are given a set of trajectories collected with K baseline policies. Each of these policies can be quite suboptim…
A Mechanism for Sample-Efficient In-Context Learning for Sparse Retrieval Tasks
Jacob Abernethy, Alekh Agarwal, Teodor V. Marinov +1
We study the phenomenon of \textit{in-context learning} (ICL) exhibited by large language models, where they can adapt to a new learning task, given a handful of labeled examples,…
Leveraging User-Triggered Supervision in Contextual Bandits
Alekh Agarwal, Claudio Gentile, Teodor V. Marinov
We study contextual bandit (CB) problems, where the user can sometimes respond with the best action in a given context. Such an interaction arises, for example, in text prediction…