1 paper
Shradha Sharma, Shweta Jain, Swapnil Dhamal
We study meritocratic fairness in budgeted combinatorial multi-armed bandits with full-bandit feedback, where a learner selects at most K arms per time step and observes only the…