1 paper
Khashayar Khosravi, Renato Paes Leme, Chara Podimata +1
We propose a model for learning with bandit feedback while accounting for deterministically evolving and unobservable states that we call Bandits with Deterministically Evolving St…