1 paper
Parsa Vares, Ãloi Durant, Jun Pang +2
Thompson Sampling (TS) and its variants are powerful Multi-Armed Bandit algorithms used to balance exploration and exploitation strategies in active learning. Yet, their probabilis…