2 citations · 2 across the 1 of their papers we have counts for
1 paper
Yangyang Zhao, Zhenyu Wang, Zhenhua Huang
Dialogue policy learning based on reinforcement learning is difficult to be applied to real users to train dialogue agents from scratch because of the high cost. User simulators, w…