1 paper
Sai Zhang, Yuwei Hu, Xiaojie Wang +1
Reinforcement learning has been applied to train the dialog systems in many works. Previous approaches divide the dialog system into multiple modules including DST (dialog state tr…