◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yaozhong Gan

10 papers hereh-index 5126 citations17 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author5
  • last author1

Across the 10 of 10 papers where every author was matched, so the position is known.

fields
  • cs.LG8
  • cs.CV1
  • cs.MA1

identity via Semantic Scholar / OpenAlex

activity
20192025
most citedStabilizing Q Learning Via Soft Mellowmax Operator

1 citations · 1 across the 9 of their papers we have counts for

collaborators
Showing 2022Show all

2 papers · 1 filter

cs.LG2022

Robust Action Gap Increasing with Clipped Advantage Learning

Zhe Zhang, Yaozhong Gan, Xiaoyang Tan

Advantage Learning (AL) seeks to increase the action gap between the optimal action and its competitors, so as to improve the robustness to estimation errors. However, the method b…

cs.LG2022

Smoothing Advantage Learning

Yaozhong Gan, Zhe Zhang, Xiaoyang Tan

Advantage learning (AL) aims to improve the robustness of value-based reinforcement learning against estimation errors with action-gap-based regularization. Unfortunately, the meth…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.