19 citations · 26 across the 5 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
stat.ML2022
Offline Reinforcement Learning for Human-Guided Human-Machine Interaction with Private Information
Zuyue Fu, Zhengling Qi, Zhuoran Yang +2
Motivated by the human-machine interaction such as training chatbots for improving customer satisfaction, we study human-guided human-machine interaction involving private informat…
cs.LG2022★ 6 cited
Offline Reinforcement Learning with Instrumental Variables in Confounded Markov Decision Processes
Zuyue Fu, Zhengling Qi, Zhaoran Wang +3
We study the offline reinforcement learning (RL) in the face of unmeasured confounders. Due to the lack of online interaction with the environment, offline RL is facing the followi…