1 paper
Xinyu Dai, Daniel Chen, Yian Qian
Dynamic decision-making under model uncertainty is central to many economic environments, yet existing bandit and reinforcement learning algorithms rely on the assumption of correc…