1 paper
Shi Feng, Nuoya Xiong, Wei Chen
In combinatorial causal bandits (CCB), the learning agent chooses a subset of variables in each round to intervene and collects feedback from the observed variables to minimize exp…