◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

My Phan

2 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • last author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG2

identity via Semantic Scholar / OpenAlex

most citedRegret Balancing for Bandit and RL Model Selection

9 citations · 9 across the 1 of their papers we have counts for

collaborators

2 papers

cs.LG2020★ 9 cited

Regret Balancing for Bandit and RL Model Selection

Yasin Abbasi-Yadkori, Aldo Pacchiano, My Phan

We consider model selection in stochastic bandit and reinforcement learning problems. Given a set of base learning algorithms, an effective model selection strategy adapts to the b…

cs.LG2019

Thompson Sampling with Approximate Inference

My Phan, Yasin Abbasi-Yadkori, Justin Domke

We study the effects of approximate inference on the performance of Thompson sampling in the k-armed bandit problems. Thompson sampling is a successful algorithm for online decis…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.