◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Paul Weng

7 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • middle author3
  • last author2

Across the 6 of 7 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.AI3
ORCID 0000-0002-2008-4569
same name
  • Paul Weng — 13 papers, h 22
  • Paul Weng — 9 papers, h 3
  • Paul Weng — 3 papers, h 1
  • Paul Weng — 1 paper, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20162023
most citedOptimizing Quantiles in Preference-based Markov Decision Processes

7 citations · 24 across the 7 of their papers we have counts for

collaborators
Showing 2016Show all

2 papers · 1 filter

cs.AI2016★ 7 cited

Optimizing Quantiles in Preference-based Markov Decision Processes

Hugo Gilbert, Paul Weng, Yan Xu

In the Markov decision process model, policies are usually evaluated by expected cumulative rewards. As this decision criterion is not always suitable, we propose in this paper an…

cs.LG2016★ 6 cited

Quantile Reinforcement Learning

Hugo Gilbert, Paul Weng

In reinforcement learning, the standard criterion to evaluate policies in a state is the expectation of (discounted) sum of rewards. However, this criterion may not always be suita…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.