31 citations · 40 across the 4 of their papers we have counts for
1 paper · 1 filter
Li Zhou, Kevin Small, Oleg Rokhlenko +1
Learning a goal-oriented dialog policy is generally performed offline with supervised learning algorithms or online with reinforcement learning (RL). Additionally, as companies acc…