1 citations · 1 across the 2 of their papers we have counts for
Showing 2019Show all
2 papers · 1 filter
cs.LG2019
HIGhER : Improving instruction following with Hindsight Generation for Experience Replay
Geoffrey Cideron, Mathieu Seurin, Florian Strub +1
Language creates a compact representation of the world and allows the description of unlimited situations and objectives through compositionality. While these characterizations may…
cs.LG2019
I'm sorry Dave, I'm afraid I can't do that, Deep Q-learning from forbidden action
Mathieu Seurin, Philippe Preux, Olivier Pietquin
The use of Reinforcement Learning (RL) is still restricted to simulation or to enhance human-operated systems through recommendations. Real-world environments (e.g. industrial robo…