397 citations · 846 across the 21 of their papers we have counts for
1 paper · 2 filters
Hao Li, Xue Yang, Zhaokai Wang +7
Many reinforcement learning environments (e.g., Minecraft) provide only sparse rewards that indicate task completion or failure with binary values. The challenge in exploration eff…