activity
20182025
most citedStraggler-Resilient Distributed Machine Learning with Dynamic Backup Workers

10 citations · 16 across the 4 of their papers we have counts for

collaborators

8 papers

eess.SY2025

Multi-Agent Reinforcement Learning for Decentralized Reservoir Management via Murmuration Intelligence

Heming Fu, Guojun Xiong, Jian Li +1

Conventional centralized water management systems face critical limitations from computational complexity and uncertainty propagation. We present MurmuRL, a novel decentralized fra…

cs.LG2024

On the Linear Speedup of Personalized Federated Reinforcement Learning with Shared Representations

Guojun Xiong, Shufan Wang, Daniel Jiang +1

Federated reinforcement learning (FedRL) enables multiple agents to collaboratively learn a policy without sharing their local trajectories collected during agent-environment inter…

cs.LG2024

DOPL: Direct Online Preference Learning for Restless Bandits with Preference Feedback

Guojun Xiong, Ujwal Dinesha, Debajoy Mukherjee +2

Restless multi-armed bandits (RMAB) has been widely used to model constrained sequential decision making problems, where the state of each restless arm evolves according to a Marko…

cs.LG2024

Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback

Guojun Xiong, Jian Li

Restless multi-armed bandits (RMAB) play a central role in modeling sequential decision making problems under an instantaneous activation constraint that at most B arms can be acti…

cs.LG202110 cited

Straggler-Resilient Distributed Machine Learning with Dynamic Backup Workers

Guojun Xiong, Gang Yan, Rahul Singh +1

With the increasing demand for large-scale training of machine learning models, consensus-based distributed optimization methods have recently been advocated as alternatives to the…

cs.NI20216 cited

Learning Augmented Index Policy for Optimal Service Placement at the Network Edge

Guojun Xiong, Rahul Singh, Jian Li

We consider the problem of service placement at the network edge, in which a decision maker has to choose between services to host at the edge to satisfy the demands of custome…