◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Ryabko

10 papers hereh-index 18982 citations95 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author2
  • middle author1
  • last author6

Across the 10 of 10 papers where every author was matched, so the position is known.

fields
  • cs.LG6
  • cs.CC1
  • cs.CR1
  • math.PR1
  • stat.ML1

identity via Semantic Scholar / OpenAlex

activity
20062013
most citedOnline Regret Bounds for Undiscounted Continuous Reinforcement Learning

44 citations · 127 across the 10 of their papers we have counts for

collaborators
Showing 2006 · cs.LGShow all

2 papers · 2 filters

cs.LG2006★ 1 cited

On Sequence Prediction for Arbitrary Measures

Daniil Ryabko, Marcus Hutter

Suppose we are given two probability measures on the set of one-way infinite finite-alphabet sequences and consider the question when one of the measures predicts the other, that i…

cs.LG2006

Asymptotic Learnability of Reinforcement Problems with Arbitrary Dependence

Daniil Ryabko, Marcus Hutter

We address the problem of reinforcement learning in which observations may exhibit an arbitrary form of stochastic dependence on past observations and actions. The task for an agen…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.