13 citations · 29 across the 9 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2022★ 1 cited
Beyond the Best: Estimating Distribution Functionals in Infinite-Armed Bandits
Yifei Wang, Tavor Baharav, Yanjun Han +2
In the infinite-armed bandit problem, each arm's average reward is sampled from an unknown distribution, and each arm can be sampled further to obtain noisy estimates of the averag…
cs.LG2021★ 3 cited
Provably Breaking the Quadratic Error Compounding Barrier in Imitation Learning, Optimally
Nived Rajaraman, Yanjun Han, Lin F. Yang +2
We study the statistical limits of Imitation Learning (IL) in episodic Markov Decision Processes (MDPs) with a state space . We focus on the known-transition setting w…
cs.LG2018
Entropy Rate Estimation for Markov Chains with Large State Space
Yanjun Han, Jiantao Jiao, Chuan-Zheng Lee +3
Estimating the entropy based on data is one of the prototypical problems in distribution property testing and estimation. For estimating the Shannon entropy of a distribution on $S…