◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Phanideep Gampa

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.IR1

identity via Semantic Scholar / OpenAlex

most citedBanditRank: Learning to Rank Using Contextual Bandits

3 citations · 3 across the 1 of their papers we have counts for

collaborators

3 papers

cs.LG2020

Object Files and Schemata: Factorizing Declarative and Procedural Knowledge in Dynamical Systems

Anirudh Goyal, Alex Lamb, Phanideep Gampa +5

Modeling a structured, dynamic environment like a video game requires keeping track of the objects and their states declarative knowledge) as well as predicting how objects behave…

cs.IR2019★ 3 cited

BanditRank: Learning to Rank Using Contextual Bandits

Phanideep Gampa, Sumio Fujita

We propose an extensible deep learning method that uses reinforcement learning to train neural networks for offline ranking in information retrieval (IR). We call our method Bandit…

cs.LG2019

A Tractable Algorithm For Finite-Horizon Continuous Reinforcement Learning

Phanideep Gampa, Sairam Satwik Kondamudi, Lakshmanan Kailasam

We consider the finite horizon continuous reinforcement learning problem. Our contribution is three-fold. First,we give a tractable algorithm based on optimistic value iteration fo…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.