1 citations · 1 across the 1 of their papers we have counts for
1 paper
Daiki Kimura, Masaki Ono, Subhajit Chaudhury +6
Deep reinforcement learning (RL) methods often require many trials before convergence, and no direct interpretability of trained policies is provided. In order to achieve fast conv…