3 papers
cs.GT2026
Learning to Strategically Acquire Resources in Competition
Safwan Hossain, Mirah Shi, Andrew Bennett +4
We consider multiple agents competing to acquire some costly divisible resource (e.g. shares of a financial asset, compute resources, etc.) over time. Leveraging a standard model f…
cs.AI2024
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
Andrew Bennett, Nathan Kallus, Miruna Oprescu +2
We study the evaluation of a policy under best- and worst-case perturbations to a Markov decision process (MDP), using transition observations from the original MDP, whether they a…
cs.LG2023
Low-Rank MDPs with Continuous Action Spaces
Andrew Bennett, Nathan Kallus, Miruna Oprescu
Low-Rank Markov Decision Processes (MDPs) have recently emerged as a promising framework within the domain of reinforcement learning (RL), as they allow for provably approximately…