103 citations · 167 across the 3 of their papers we have counts for
Showing stat.MLShow all
3 papers · 1 filter
stat.ML2017
Action-depedent Control Variates for Policy Optimization via Stein's Identity
Hao Liu, Yihao Feng, Yi Mao +3
Policy gradient methods have achieved remarkable successes in solving challenging reinforcement learning problems. However, it still often suffers from the large variance issue on…
stat.ML2017
Sequence Modeling via Segmentations
Chong Wang, Yining Wang, Po-Sen Huang +3
Segmental structure is a common pattern in many types of sequences such as phrases in human languages. In this paper, we present a probabilistic model for sequences via their segme…
stat.ML2016
Exact Exponent in Optimal Rates for Crowdsourcing
Chao Gao, Yu Lu, Dengyong Zhou
In many machine learning applications, crowdsourcing has become the primary means for label collection. In this paper, we study the optimal error rate for aggregating labels provid…