◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Anurag Koul

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG3

identity via Semantic Scholar / OpenAlex

activity
20182022
collaborators

3 papers

cs.LG2022

Offline Policy Comparison with Confidence: Benchmarks and Baselines

Anurag Koul, Mariano Phielipp, Alan Fern

Decision makers often wish to use offline historical data to compare sequential-action policies at various world states. Importantly, computational tools should produce confidence…

cs.LG2020

Dream and Search to Control: Latent Space Planning for Continuous Control

Anurag Koul, Varun V. Kumar, Alan Fern +1

Learning and planning with latent space dynamics has been shown to be useful for sample efficiency in model-based reinforcement learning (MBRL) for discrete and continuous control…

cs.LG2018

Learning Finite State Representations of Recurrent Policy Networks

Anurag Koul, Sam Greydanus, Alan Fern

Recurrent neural networks (RNNs) are an effective representation of control policies for a wide range of reinforcement and imitation learning problems. RNN policies, however, are p…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.