2 citations · 2 across the 1 of their papers we have counts for
1 paper
Elita A. Lobo, Cyrus Cousins, Yair Zick +1
In reinforcement learning, robust policies for high-stakes decision-making problems with limited data are usually computed by optimizing the \emph{percentile criterion}. The percen…