1 paper · 1 filter
Alberta Longhini, David Emukpere, Jean-Michel Renders +1
We address the problem of fine-tuning pre-trained generative policies with reinforcement learning (RL) while preserving the multimodality of their action distributions. Existing me…