1 paper
Qihan Guo, Siwei Wang, Jun Zhu
We study an extension of standard bandit problem in which there are R layers of experts. Multi-layered experts make selections layer by layer and only the experts in the last layer…