1 paper
Taha Shieenavaz, Shabnam Zareshahraki, Loris Nanni
Replay-free parallelized Q-learning removes the large experience replay buffers and target networks used by conventional deep Q-learning, but the role of network architecture in th…