◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ian Osband

5 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author1

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • stat.ML3
  • cs.LG2
same name
  • Ian Osband — 16 papers, h 28

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20142023
most citedOn Lower Bounds for Regret in Reinforcement Learning

49 citations · 101 across the 5 of their papers we have counts for

collaborators
Showing stat.MLShow all

3 papers · 1 filter

stat.ML2016★ 49 cited

On Lower Bounds for Regret in Reinforcement Learning

Ian Osband, Benjamin Van Roy

This is a brief technical note to clarify the state of lower bounds on regret for reinforcement learning. In particular, this paper: - Reproduces a lower bound on regret for reinfo…

stat.ML2016★ 8 cited

Posterior Sampling for Reinforcement Learning Without Episodes

Ian Osband, Benjamin Van Roy

This is a brief technical note to clarify some of the issues with applying the application of the algorithm posterior sampling for reinforcement learning (PSRL) in environments wit…

stat.ML2014★ 43 cited

Model-based Reinforcement Learning and the Eluder Dimension

Ian Osband, Benjamin Van Roy

We consider the problem of learning to optimize an unknown Markov decision process (MDP). We show that, if the MDP can be parameterized within some known function class, we can obt…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.