Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Provably Efficient and Agile Randomized Q-Learning
He Wang, Xingyu Xu, Yuejie Chi
While Bayesian-based exploration often demonstrates superior empirical performance compared to bonus-based methods in model-based reinforcement learning (RL), its theoretical under…
cs.LG2025
Communication-Efficient Federated Optimization over Semi-Decentralized Networks
He Wang, Yuejie Chi
In large-scale federated and decentralized learning, communication efficiency is one of the most challenging bottlenecks. While gossip communication -- where agents can exchange in…
cs.LG2024
Sample Complexity of Offline Distributionally Robust Linear Markov Decision Processes
He Wang, Laixi Shi, Yuejie Chi
In offline reinforcement learning (RL), the absence of active exploration calls for attention on the model robustness to tackle the sim-to-real gap, where the discrepancy between t…