3 papers
cs.LG2026
Shared Actors Need Not Share Critics: Effects of Value Mismatch in Parallel Reinforcement Learning
Zhenya Liu, Yang Meng, Zhuokai Zhao +2
When a single policy is trained in parallel across multiple environments of the same task, such as procedurally generated levels, randomized dynamics, or curricula, implementations…
cs.LG2026
Active Curriculum Refinement for Reinforcement Learning
Zhenya Liu, Yuxin Chen
In many reinforcement learning (RL) domains, environments are connected by prerequisite relations, such as difficulty-increasing edits or parameter increments, which induce a direc…
cs.LG2026
Rethinking Transfer in Continual Learning: A Replay-Based Realisation
Yang Meng, Zhenya Liu, Zhuokai Zhao +1
Continual learning studies how deployed language models can continually acquire new tasks without expensive retraining from scratch. Existing methods, whether rehearsal-based (repl…