1 paper
Tamojeet Roychowdhury, Kota Srinivas Reddy, Krishna P Jagannathan +1
We focus on the problem of best-arm identification in a stochastic multi-arm bandit with temporally decreasing variances for the arms' rewards. We model arm rewards as Gaussian ran…