1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 1 cited
SaFormer: A Conditional Sequence Modeling Approach to Offline Safe Reinforcement Learning
Qin Zhang, Linrui Zhang, Haoran Xu +6
Offline safe RL is of great practical relevance for deploying agents in real-world applications. However, acquiring constraint-satisfying policies from the fixed dataset is non-tri…
cs.LG2021
Probability Density Estimation Based Imitation Learning
Yang Liu, Yongzhe Chang, Shilei Jiang +3
Imitation Learning (IL) is an effective learning paradigm exploiting the interactions between agents and environments. It does not require explicit reward signals and instead tries…