2 citations · 2 across the 1 of their papers we have counts for
1 paper
Jiafei Lyu, Aicheng Gong, Le Wan +2
We present state advantage weighting for offline reinforcement learning (RL). In contrast to action advantage A(s,a) that we commonly adopt in QSA learning, we leverage state adv…