3 papers
cs.LG2026
Conditional Sequence Modeling for Safe Reinforcement Learning
Wensong Bai, Chao Zhang, Qihang Xu +3
Offline safe reinforcement learning (RL) aims to learn policies from a fixed dataset while maximizing performance under cumulative cost constraints. In practice, deployment require…
cs.LG2023
Towards Optimal Randomized Strategies in Adversarial Example Game
Jiahao Xie, Chao Zhang, Weijie Liu +2
The vulnerability of deep neural network models to adversarial example attacks is a practical challenge in many artificial intelligence applications. A recent line of work shows th…
cs.LG2023
PACER: A Fully Push-forward-based Distributional Reinforcement Learning Algorithm
Wensong Bai, Chao Zhang, Yichao Fu +3
In this paper, we propose the first fully push-forward-based distributional reinforcement learning algorithm, named PACER, which consists of a distributional critic, a stochastic a…