3 citations · 12 across the 16 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Off-Policy Primal-Dual Safe Reinforcement Learning
Zifan Wu, Bo Tang, Qian Lin +5
Primal-dual safe RL methods commonly perform iterations between the primal update of the policy and the dual update of the Lagrange Multiplier. Such a training paradigm is highly s…
cs.LG2016★ 3 cited
Memory Visualization for Gated Recurrent Neural Networks in Speech Recognition
Zhiyuan Tang, Ying Shi, Dong Wang +2
Recurrent neural networks (RNNs) have shown clear superiority in sequence modeling, particularly the ones with gated units, such as long short-term memory (LSTM) and gated recurren…