2 papers
cs.LG2021
Off-Dynamics Inverse Reinforcement Learning from Hetero-Domain
Yachen Kang, Jinxin Liu, Xin Cao +1
We propose an approach for inverse reinforcement learning from hetero-domain which learns a reward function in the simulator, drawing on the demonstrations from the real world. The…
cs.SI2018
Improved Online Wilson Score Interval Method for Community Answer Quality Ranking
Xin Cao
In this paper, a fast and easy-to-deploy method with a strong interpretability for community answer quality ranking is proposed. This method is improved based on the Wilson score i…