31 citations · 31 across the 1 of their papers we have counts for
1 paper · 1 filter
Li Zhou, Kevin Small, Oleg Rokhlenko +1
Learning a goal-oriented dialog policy is generally performed offline with supervised learning algorithms or online with reinforcement learning (RL). Additionally, as companies acc…