◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Leif Döring

6 papers hereh-index 334 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • last author5

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • math.OC1
  • stat.ML1

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2025

ADDQ: Adaptive Distributional Double Q-Learning

Leif Döring, Benedikt Wille, Maximilian Birr +2

Bias problems in the estimation of Q-values are a well-known obstacle that slows down convergence of Q-learning and actor-critic methods. One of the reasons of the success of m…

cs.LG2025

Almost sure convergence rates of stochastic gradient methods under gradient domination

Simon Weissmann, Sara Klein, Waïss Azizian +1

Stochastic gradient methods are among the most important algorithms in training machine learning problems. While classical assumptions such as strong convexity allow a simple analy…

cs.LG2025

Clustered KL-barycenter design for policy evaluation

Simon Weissmann, Till Freihaut, Claire Vernade +2

In the context of stochastic bandit models, this article examines how to design sample-efficient behavior policies for the importance sampling evaluation of multiple target policie…

cs.LG2024

Structure Matters: Dynamic Policy Gradient

Sara Klein, Xiangyuan Zhang, Tamer Başar +2

In this work, we study I^3-discounted infinite-horizon tabular Markov decision processes (MDPs) and introduce a framework called dynamic policy gradient (DynPG). The framework dir…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.