2 citations · 2 across the 1 of their papers we have counts for
1 paper
Anish Agarwal, Abdullah Alomar, Varkey Alumootil +4
We consider offline reinforcement learning (RL) with heterogeneous agents under severe data scarcity, i.e., we only observe a single historical trajectory for every agent under an…