1 paper
Ishan Durugkar, Steven Hansen, Stephen Spencer +1
This paper deals with the problem of learning a skill-conditioned policy that acts meaningfully in the absence of a reward signal. Mutual information based objectives have shown so…