8 citations · 14 across the 2 of their papers we have counts for
3 papers
cs.LG2021★ 6 cited
Near-optimal Representation Learning for Linear Bandits and Linear RL
Jiachen Hu, Xiaoyu Chen, Chi Jin +2
This paper studies representation learning for multi-task linear bandits and multi-task episodic RL with linear value function approximation. We first consider the setting where we…
cs.LG2020
Efficient Reinforcement Learning in Factored MDPs with Application to Constrained RL
Xiaoyu Chen, Jiachen Hu, Lihong Li +1
Reinforcement learning (RL) in episodic, factored Markov decision processes (FMDPs) is studied. We propose an algorithm called FMDP-BF, which leverages the factorization structure…
cs.LG2019★ 8 cited
Distributed Bandit Learning: Near-Optimal Regret with Efficient Communication
Yuanhao Wang, Jiachen Hu, Xiaoyu Chen +1
We study the problem of regret minimization for distributed bandits learning, in which agents work collaboratively to minimize their total regret under the coordination of a ce…