7 citations · 7 across the 1 of their papers we have counts for
1 paper
Kirthevasan Kandasamy, Yoram Bachrach, Ryota Tomioka +2
We study reinforcement learning of chatbots with recurrent neural network architectures when the rewards are noisy and expensive to obtain. For instance, a chatbot used in automate…