Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning
Chen-Xiao Gao, Chenyang Wu, Mingjun Cao +3
Behavior regularization, which constrains the policy to stay close to some behavior policy, is widely used in offline reinforcement learning (RL) to manage the risk of hazardous ex…
cs.LG2024
Reinforced In-Context Black-Box Optimization
Lei Song, Chenxiao Gao, Ke Xue +5
Black-Box Optimization (BBO) has found successful applications in many fields of science and engineering. Recently, there has been a growing interest in meta-learning particular co…