2 papers
cs.LG2024
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
Yassine Chemingui, Aryan Deshwal, Honghao Wei +2
Offline safe reinforcement learning (OSRL) involves learning a decision-making policy to maximize rewards from a fixed batch of training data to satisfy pre-defined safety constrai…
cs.LG2024
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
Keru Chen, Honghao Wei, Zhigang Deng +1
The high costs and risks involved in extensive environment interactions hinder the practical application of current online safe reinforcement learning (RL) methods. While offline s…