117 citations · 152 across the 13 of their papers we have counts for
Showing 2018Show all
3 papers · 1 filter
cs.CV2018
Anchor Box Optimization for Object Detection
Yuanyi Zhong, Jianfeng Wang, Jian Peng +1
In this paper, we propose a general approach to optimize anchor boxes for object detection. Nowadays, anchor boxes are widely adopted in state-of-the-art detection frameworks. Howe…
stat.ML2018
Learning Self-Imitating Diverse Policies
Tanmay Gangwani, Qiang Liu, Jian Peng
The success of popular algorithms for deep reinforcement learning, such as policy-gradients and Q-learning, relies heavily on the availability of an informative reward signal at ea…
cs.LG2018
Learning to Explore with Meta-Policy Gradient
Tianbing Xu, Qiang Liu, Liang Zhao +1
The performance of off-policy learning, including deep Q-learning and deep deterministic policy gradient (DDPG), critically depends on the choice of the exploration policy. Existin…