◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Guanghui Lan

3 papers hereh-index 27 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author2

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • math.OC1
same name
  • Guanghui Lan — 20 papers, h 39
  • Guanghui Lan — 7 papers
  • Guanghui Lan — 4 papers
  • Guanghui Lan — 3 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

math.OC2026

Value Mirror Descent for Reinforcement Learning

Zhichao Jia, Guanghui Lan

Value iteration-type methods have been extensively studied for computing a nearly optimal value function in reinforcement learning (RL). Under a generative sampling model, these me…

cs.LG2026

Actor-Accelerated Policy Dual Averaging for Reinforcement Learning in Continuous Action Spaces

Ji Gao, Caleb Ju, Guanghui Lan +1

Policy Dual Averaging (PDA) offers a principled Policy Mirror Descent (PMD) framework that more naturally admits value function approximation than standard PMD, enabling the use of…

cs.LG2024

Strongly-polynomial time and validation analysis of policy gradient methods

Caleb Ju, Guanghui Lan

This paper proposes a novel termination criterion, termed the advantage gap function, for finite state and action Markov decision processes (MDP) and reinforcement learning (RL). B…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.