8 citations · 17 across the 8 of their papers we have counts for
3 papers · 1 filter
Bootstrapped Q-learning with Context Relevant Observation Pruning to Generalize in Text-based Games
Subhajit Chaudhury, Daiki Kimura, Kartik Talamadupula +3
We show that Reinforcement Learning (RL) methods for solving Text-Based Games (TBGs) often fail to generalize on unseen games, especially in small data regimes. To address this iss…
Injective State-Image Mapping facilitates Visual Adversarial Imitation Learning
Subhajit Chaudhury, Daiki Kimura, Asim Munawar +1
The growing use of virtual autonomous agents in applications like games and entertainment demands better control policies for natural-looking movements and actions. Unlike the conv…
Internal Model from Observations for Reward Shaping
Daiki Kimura, Subhajit Chaudhury, Ryuki Tachibana +1
Reinforcement learning methods require careful design involving a reward function to obtain the desired action policy for a given task. In the absence of hand-crafted reward functi…