◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Mengdi Wang

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • stat.ML2
  • cs.LG1
  • math.OC1
ORCID 0000-0002-2101-9507
same name
  • Mengdi Wang — 3 papers
  • Mengdi Wang — 1 paper
  • Mengdi Wang — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20142019
most citedStochastic Primal-Dual Methods and Sample Complexity of Reinforcement Learning

42 citations · 112 across the 4 of their papers we have counts for

collaborators

4 papers

cs.LG2019★ 40 cited

Reinforcement Learning in Feature Space: Matrix Bandit, Kernels, and Regret Bound

Lin F. Yang, Mengdi Wang

Exploration in reinforcement learning (RL) suffers from the curse of dimensionality when the state-action space is large. A common practice is to parameterize the high-dimensional…

stat.ML2016★ 42 cited

Stochastic Primal-Dual Methods and Sample Complexity of Reinforcement Learning

Yichen Chen, Mengdi Wang

We study the online estimation of the optimal policy of a Markov decision process (MDP). We propose a class of Stochastic Primal-Dual (SPD) methods which exploit the inherent minim…

math.OC2016★ 24 cited

Accelerating Stochastic Composition Optimization

Mengdi Wang, Ji Liu, Ethan X. Fang

Consider the stochastic composition optimization problem where the objective is a composition of two expected-value functions. We propose a new stochastic first-order method, namel…

stat.ML2014★ 6 cited

Stochastic Compositional Gradient Descent: Algorithms for Minimizing Compositions of Expected-Value Functions

Mengdi Wang, Ethan X. Fang, Han Liu

Classical stochastic gradient methods are well suited for minimizing expected-value objective functions. However, they do not apply to the minimization of a nonlinear function invo…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.