12 citations · 13 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2019
Multi Pseudo Q-learning Based Deterministic Policy Gradient for Tracking Control of Autonomous Underwater Vehicles
Wenjie Shi, Shiji Song, Cheng Wu +1
This paper investigates trajectory tracking problem for a class of underactuated autonomous underwater vehicles (AUVs) with unknown dynamics and constrained inputs. Different from…
cs.LG2019
Soft Policy Gradient Method for Maximum Entropy Deep Reinforcement Learning
Wenjie Shi, Shiji Song, Cheng Wu
Maximum entropy deep reinforcement learning (RL) methods have been demonstrated on a range of challenging continuous tasks. However, existing methods either suffer from severe inst…