3 papers
cs.LG2024
Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin Representation
Wanpeng Zhang, Yilin Li, Boyu Yang +1
In real-world scenarios, the application of reinforcement learning is significantly challenged by complex non-stationarity. Most existing methods attempt to model changes in the en…
cs.AI2024
AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback
Wanpeng Zhang, Zongqing Lu
Large Language Models (LLMs) have demonstrated significant success across various domains. However, their application in complex decision-making tasks frequently necessitates intri…
cs.LG2024
MBDP: A Model-based Approach to Achieve both Robustness and Sample Efficiency via Double Dropout Planning
Wanpeng Zhang, Xi Xiao, Yao Yao +2
Model-based reinforcement learning is a widely accepted solution for solving excessive sample demands. However, the predictions of the dynamics models are often not accurate enough…