1 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.LG2022★ 1 cited
Selectively Contextual Bandits
Claudia Roberts, Maria Dimakopoulou, Qifeng Qiao +2
Contextual bandits are widely used in industrial personalization systems. These online learning frameworks learn a treatment assignment policy in the presence of treatment effects…
cs.LG2013★ 1 cited
Behavior Pattern Recognition using A New Representation Model
Qifeng Qiao, Peter A. Beling
We study the use of inverse reinforcement learning (IRL) as a tool for the recognition of agents' behavior on the basis of observation of their sequential decision behavior interac…
cs.LG2012
Inverse Reinforcement Learning with Gaussian Process
Qifeng Qiao, Peter A. Beling
We present new algorithms for inverse reinforcement learning (IRL, or inverse optimal control) in convex optimization settings. We argue that finite-space IRL can be posed as a con…