4 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CL2022★ 4 cited
CHAI: A CHatbot AI for Task-Oriented Dialogue with Offline Reinforcement Learning
Siddharth Verma, Justin Fu, Mengjiao Yang +1
Conventionally, generation of natural language for dialogue agents may be viewed as a statistical learning problem: determine the patterns in human-provided data and generate appro…
cs.LG2020★ 3 cited
Continual Learning of Control Primitives: Skill Discovery via Reset-Games
Kelvin Xu, Siddharth Verma, Chelsea Finn +1
Reinforcement learning has the potential to automate the acquisition of behavior in complex settings, but in order for it to be successfully deployed, a number of practical challen…
cs.LG2019
Fast Online "Next Best Offers" using Deep Learning
Rekha Singhal, Gautam Shroff, Mukund Kumar +5
In this paper, we present iPrescribe, a scalable low-latency architecture for recommending 'next-best-offers' in an online setting. The paper presents the design of iPrescribe and…