3 citations · 6 across the 4 of their papers we have counts for
1 paper · 1 filter
Raymond Feng, Jesse Geneson, Andrew Lee +1
We determine sharp bounds on the price of bandit feedback for several variants of the mistake-bound model. The first part of the paper presents bounds on the r-input weak reinfor…