2 papers
cs.LG2026
The Price of Decentralization in Top- Arm Identification
Larissa Xu, Jasmine Nguyen, William Chang
Cooperative teams often need to agree on the best few options rather than simply accumulate reward, and they must do so while each member sees only a fragment of the team's collect…
cs.LG2026
Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry
Larissa Xu, King Bi, William Chang
We study decentralized multi-player reinforcement learning in episodic tabular Markov decision processes (MDPs) under three forms of information asymmetry: (A) unobserved actions w…