5 citations · 5 across the 1 of their papers we have counts for
1 paper · 1 filter
Aakash Maroti
ε-greedy is a policy used to balance exploration and exploitation in many reinforcement learning setting. In cases where the agent uses some on-policy algorithm to lear…