1 paper
Golnaz Mesbahi, Parham Mohammad Panahi, Olya Mastikhina +3
In continual RL we want agents capable of never-ending learning, and yet our evaluation methodologies do not reflect this. The standard practice in RL is to assume unfettered acces…