Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Continual Knowledge Adaptation for Reinforcement Learning
Jinwu Hu, Zihao Lian, Zhiquan Wen +5
Reinforcement Learning enables agents to learn optimal behaviors through interactions with environments. However, real-world environments are typically non-stationary, requiring ag…
cs.AI2025
Efficient Dynamic Ensembling for Multiple LLM Experts
Jinwu Hu, Yufeng Wang, Shuhai Zhang +5
LLMs have demonstrated impressive performance across various language tasks. However, the strengths of LLMs can vary due to different architectures, model sizes, areas of training…