3 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 3 cited
Multimodal Reward Shaping for Efficient Exploration in Reinforcement Learning
Mingqi Yuan, Mon-on Pun, Dong Wang +2
Maintaining the long-term exploration capability of the agent remains one of the critical challenges in deep reinforcement learning. A representative solution is to leverage reward…
cs.OS2020★ 3 cited
Fairness-Oriented User Scheduling for Bursty Downlink Transmission Using Multi-Agent Reinforcement Learning
Mingqi Yuan, Qi Cao, Man-on Pun +1
In this work, we develop practical user scheduling algorithms for downlink bursty traffic with emphasis on user fairness. In contrast to the conventional scheduling algorithms that…