298 citations · 671 across the 40 of their papers we have counts for
3 papers · 1 filter
Learning Natural Language Generation from Scratch
Alice Martin Donati, Guillaume Quispe, Charles Ollion +3
This paper introduces TRUncated ReinForcement Learning for Language (TrufLL), an original ap-proach to train conditional language models from scratch by only using reinforcement le…
Scaling up Mean Field Games with Online Mirror Descent
Julien Perolat, Sarah Perrin, Romuald Elie +5
We address scaling up equilibrium computation in Mean Field Games (MFGs) using Online Mirror Descent (OMD). We show that continuous-time OMD provably converges to a Nash equilibriu…
Countering Language Drift with Seeded Iterated Learning
Yuchen Lu, Soumye Singhal, Florian Strub +2
Pretraining on human corpus and then finetuning in a simulator has become a standard pipeline for training a goal-oriented dialogue agent. Nevertheless, as soon as the agents are f…