2 papers
cs.LG2026
LexiSafe: Offline Safe Reinforcement Learning with Lexicographic Safety-Reward Hierarchy
Hsin-Jung Yang, Zhanhong Jiang, Prajwal Koirala +3
Offline safe reinforcement learning (RL) is increasingly important for cyber-physical systems (CPS), where safety violations during training are unacceptable and only pre-collected…
cs.LG2026
Directional Concentration Uncertainty: A representational approach to uncertainty quantification for generative models
Souradeep Chattopadhyay, Brendan Kennedy, Sai Munikoti +2
In the critical task of making generative models trustworthy and robust, methods for Uncertainty Quantification (UQ) have begun to show encouraging potential. However, many of thes…