9 citations · 22 across the 5 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2020
Adaptive Dialog Policy Learning with Hindsight and User Modeling
Yan Cao, Keting Lu, Xiaoping Chen +1
Reinforcement learning methods have been used to compute dialog policies from language-based interaction experiences. Efficiency is of particular importance in dialog policy learni…
cs.AI2020★ 2 cited
Learning and Reasoning for Robot Dialog and Navigation Tasks
Keting Lu, Shiqi Zhang, Peter Stone +1
Reinforcement learning and probabilistic reasoning algorithms aim at learning from interaction experiences and reasoning with probabilistic contextual knowledge respectively. In th…
cs.AI2012★ 9 cited
FHHOP: A Factored Hybrid Heuristic Online Planning Algorithm for Large POMDPs
Zhongzhang Zhang, Xiaoping Chen
Planning in partially observable Markov decision processes (POMDPs) remains a challenging topic in the artificial intelligence community, in spite of recent impressive progress in…