Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated Value Dynamics
Bozhou Chen, Yongyi Wang, Hanyu Liu +2
Deep Q-learning (DQL) has achieved remarkable empirical success in reinforcement learning, yet its training process remains notoriously unstable. Existing studies often attribute i…
cs.LG2026
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
Yongyi Wang, Hanyu Liu, Lingfeng Li +6
Decision Transformer (DT) formulates offline reinforcement learning as autoregressive sequence modeling, achieving promising results by predicting actions from a sequence of Return…
cs.LG2026
Pareto-guided Pipeline for Distilling Featherweight AI Agents in Mobile MOBA Games
Xionghui Yang, Bozhou Chen, Yunlong Lu +8
Recent advances in game AI have demonstrated the feasibility of training agents that surpass top-tier human professionals in complex environments such as Honor of Kings (HoK), a le…