1 paper
Tristan Deleu, Padideh Nouri, Yoshua Bengio +1
Recent progress in generative modeling has highlighted the importance of Reinforcement Learning (RL) for fine-tuning, with KL-regularized methods in particular proving to be highly…