8 citations · 8 across the 1 of their papers we have counts for
2 papers
cs.LG2020
Online Learning in Unknown Markov Games
Yi Tian, Yuanhao Wang, Tiancheng Yu +1
We study online learning in unknown Markov games, a problem that arises in episodic multi-agent reinforcement learning where the actions of the opponents are unobservable. We show…
cs.LG2019★ 8 cited
Distributed Bandit Learning: Near-Optimal Regret with Efficient Communication
Yuanhao Wang, Jiachen Hu, Xiaoyu Chen +1
We study the problem of regret minimization for distributed bandits learning, in which agents work collaboratively to minimize their total regret under the coordination of a ce…