3 papers
cs.LG2026
SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy
Yassine Chemingui, Chenhua Fan, Honghao Wei +1
Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a criterion that often fails to detect rare but catastrophic tail events. To overcome t…
cs.LG2025
Online Optimization for Offline Safe Reinforcement Learning
Yassine Chemingui, Aryan Deshwal, Alan Fern +2
We study the problem of Offline Safe Reinforcement Learning (OSRL), where the goal is to learn a reward-maximizing policy from fixed data under a cumulative cost constraint. We pro…
cs.LG2024
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
Yassine Chemingui, Aryan Deshwal, Honghao Wei +2
Offline safe reinforcement learning (OSRL) involves learning a decision-making policy to maximize rewards from a fixed batch of training data to satisfy pre-defined safety constrai…