◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

M. Katehakis

5 papers hereh-index 262.7k citations126 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author1

Across the 4 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • stat.ML2
  • math.OC1

identity via Semantic Scholar / OpenAlex

activity
20152023
most citedNormal Bandits of Unknown Means and Variances: Asymptotic Optimality, Finite Horizon Regret Bounds, and a Solution to an Open Problem

8 citations · 8 across the 3 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2019

Accelerating the Computation of UCB and Related Indices for Reinforcement Learning

Wesley Cowan, Michael N. Katehakis, Daniel Pirutinsky

In this paper we derive an efficient method for computing the indices associated with an asymptotically optimal upper confidence bound algorithm (MDP-UCB) of Burnetas and Katehakis…

cs.LG2019

Reinforcement Learning: a Comparison of UCB Versus Alternative Adaptive Policies

Wesley Cowan, Michael N. Katehakis, Daniel Pirutinsky

In this paper we consider the basic version of Reinforcement Learning (RL) that involves computing optimal data driven (adaptive) policies for Markovian decision process with unkno…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.