8 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.LG2018★ 8 cited
Likelihood Quantile Networks for Coordinating Multi-Agent Reinforcement Learning
Xueguang Lyu, Christopher Amato
When multiple agents learn in a decentralized manner, the environment appears non-stationary from the perspective of an individual agent due to the exploration and learning of the…
stat.ML2017
On Adaptive Estimation for Dynamic Bernoulli Bandits
Xue Lu, Niall Adams, Nikolas Kantas
The multi-armed bandit (MAB) problem is a classic example of the exploration-exploitation dilemma. It is concerned with maximising the total rewards for a gambler by sequentially p…