2 papers
cs.LG2026
Provably Efficient and Agile Randomized Q-Learning
He Wang, Xingyu Xu, Yuejie Chi
While Bayesian-based exploration often demonstrates superior empirical performance compared to bonus-based methods in model-based reinforcement learning (RL), its theoretical under…
cs.LG2025
Communication-Efficient Federated Optimization over Semi-Decentralized Networks
He Wang, Yuejie Chi
In large-scale federated and decentralized learning, communication efficiency is one of the most challenging bottlenecks. While gossip communication -- where agents can exchange in…