13 citations · 29 across the 11 of their papers we have counts for
Showing 2021Show all
2 papers · 1 filter
cs.LG2021★ 3 cited
Provably Breaking the Quadratic Error Compounding Barrier in Imitation Learning, Optimally
Nived Rajaraman, Yanjun Han, Lin F. Yang +2
We study the statistical limits of Imitation Learning (IL) in episodic Markov Decision Processes (MDPs) with a state space . We focus on the known-transition setting w…
stat.ML2021★ 6 cited
Adversarial Combinatorial Bandits with General Non-linear Reward Functions
Xi Chen, Yanjun Han, Yining Wang
In this paper we study the adversarial combinatorial bandit with a known non-linear reward function, extending existing work on adversarial linear combinatorial bandit. {The advers…