2 papers
cs.CL2026
Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing
Jie Jiang, Xing Sun, Ruotian Chen +2
Speculative decoding accelerates LLM inference by having a lightweight draft model propose speculative windows of candidate tokens for parallel verification by a larger target mode…
cs.AI2025
BuildingGym: An open-source toolbox for AI-based building energy management using reinforcement learning
Xilei Dai, Ruotian Chen, Songze Guan +2
Reinforcement learning (RL) has proven effective for AI-based building energy management. However, there is a lack of flexible framework to implement RL across various control prob…