◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yue Wang

5 papers hereh-index 9376 citations18 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author1

Across the 4 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • eess.SY1
same name
  • Yue Wang — 75 papers, h 32
  • Yue Wang — 26 papers, h 22
  • Yue Wang — 21 papers, h 6
  • Yue Wang — 20 papers, h 12
  • Yue Wang — 17 papers, h 16
  • Yue Wang — 15 papers, h 12

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedPolicy Gradient Method For Robust Reinforcement Learning

11 citations · 22 across the 5 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2023★ 3 cited

Model-Free Robust Average-Reward Reinforcement Learning

Yue Wang, Alvaro Velasquez, George Atia +2

Robust Markov decision processes (MDPs) address the challenge of model uncertainty by optimizing the worst-case performance over an uncertainty set of MDPs. In this paper, we focus…

cs.LG2023

Achieving the Asymptotically Optimal Sample Complexity of Offline Reinforcement Learning: A DRO-Based Approach

Yue Wang, Jinjun Xiong, Shaofeng Zou

Offline reinforcement learning aims to learn from pre-collected datasets without active exploration. This problem faces significant challenges, including limited data availability…

cs.LG2022★ 1 cited

Robust Constrained Reinforcement Learning

Yue Wang, Fei Miao, Shaofeng Zou

Constrained reinforcement learning is to maximize the expected reward subject to constraints on utilities/costs. However, the training environment may not be the same as the test o…

cs.LG2022★ 11 cited

Policy Gradient Method For Robust Reinforcement Learning

Yue Wang, Shaofeng Zou

This paper develops the first policy gradient method with global optimality guarantee and complexity analysis for robust reinforcement learning under model mismatch. Robust reinfor…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.