5 citations · 5 across the 1 of their papers we have counts for
1 paper
Aakash Maroti
ε-greedy is a policy used to balance exploration and exploitation in many reinforcement learning setting. In cases where the agent uses some on-policy algorithm to lear…