2 papers
cs.CL2021
Studying word order through iterative shuffling
Nikolay Malkin, Sameera Lanka, Pranav Goel +1
As neural language models approach human performance on NLP benchmark tasks, their advances are widely seen as evidence of an increasingly complex understanding of syntax. This vie…
cs.LG2018
ARCHER: Aggressive Rewards to Counter bias in Hindsight Experience Replay
Sameera Lanka, Tianfu Wu
Experience replay is an important technique for addressing sample-inefficiency in deep reinforcement learning (RL), but faces difficulty in learning from binary and sparse rewards…