1 paper
Ahmed Hendawy, Henrik Metternich, Théo Vincent +3
The use of target networks is a popular approach for estimating value functions in deep Reinforcement Learning (RL). While effective, the target network remains a compromise soluti…