3 citations · 7 across the 5 of their papers we have counts for
Showing 2020 · cs.LGShow all
2 papers · 2 filters
cs.LG2020
Wasserstein Distance guided Adversarial Imitation Learning with Reward Shape Exploration
Ming Zhang, Yawei Wang, Xiaoteng Ma +4
The generative adversarial imitation learning (GAIL) has provided an adversarial learning framework for imitating expert policy from demonstrations in high-dimensional continuous t…
cs.LG2020
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
Xiaoteng Ma, Junyao Chen, Li Xia +3
We present Distributional Soft Actor-Critic (DSAC), a distributional reinforcement learning (RL) algorithm that combines the strengths of distributional information of accumulated…