16 citations · 17 across the 6 of their papers we have counts for
1 paper · 1 filter
Christian Geishauser, Songbo Hu, Hsien-chin Lin +5
The dialogue management component of a task-oriented dialogue system is typically optimised via reinforcement learning (RL). Optimisation via RL is highly susceptible to sample ine…