3 citations · 3 across the 1 of their papers we have counts for
3 papers
cs.LG2020
Object Files and Schemata: Factorizing Declarative and Procedural Knowledge in Dynamical Systems
Anirudh Goyal, Alex Lamb, Phanideep Gampa +5
Modeling a structured, dynamic environment like a video game requires keeping track of the objects and their states declarative knowledge) as well as predicting how objects behave…
cs.IR2019★ 3 cited
BanditRank: Learning to Rank Using Contextual Bandits
Phanideep Gampa, Sumio Fujita
We propose an extensible deep learning method that uses reinforcement learning to train neural networks for offline ranking in information retrieval (IR). We call our method Bandit…
cs.LG2019
A Tractable Algorithm For Finite-Horizon Continuous Reinforcement Learning
Phanideep Gampa, Sairam Satwik Kondamudi, Lakshmanan Kailasam
We consider the finite horizon continuous reinforcement learning problem. Our contribution is three-fold. First,we give a tractable algorithm based on optimistic value iteration fo…