1 citations · 1 across the 1 of their papers we have counts for
1 paper
Motoki Omura, Takayuki Osa, Yusuke Mukuta +1
In deep reinforcement learning, estimating the value function to evaluate the quality of states and actions is essential. The value function is often trained using the least square…