17 citations · 21 across the 3 of their papers we have counts for
3 papers
stat.ML2016★ 17 cited
A Batch, Off-Policy, Actor-Critic Algorithm for Optimizing the Average Reward
S. A. Murphy, Y. Deng, E. B. Laber +3
We develop an off-policy actor-critic algorithm for learning an optimal policy from a training set composed of data from multiple individuals. This algorithm is developed with a vi…
cs.LG2012
Small Sample Inference for Generalization Error in Classification Using the CUD Bound
Eric B. Laber, Susan A. Murphy
Confidence measures for the generalization error are crucial when small training samples are used to construct classifiers. A common approach is to estimate the generalization erro…
cs.LG2012★ 4 cited
Active Learning for Developing Personalized Treatment
Kun Deng, Joelle Pineau, Susan A. Murphy
The personalization of treatment via bio-markers and other risk categories has drawn increasing interest among clinical scientists. Personalized treatment strategies can be learned…