1 paper
Nicklas Werge, Yi-Shan Wu, Manuel Haussmann +2
Ensembles are ubiquitous in off-policy actor-critic learning, yet their efficacy depends critically on how they are aggregated. Current methods typically rely on static rules or ta…