◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

P. Vamplew

11 papers hereh-index 285.7k citations142 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author6
  • last author2

Across the 11 of 11 papers where every author was matched, so the position is known.

fields
  • cs.LG9
  • cs.AI1
  • cs.CY1

identity via Semantic Scholar / OpenAlex

works on
anomaly injection 1benchmark 1out-of-distribution detection 1reinforcement learning 1temporal dynamics 1

From the 1 of 11 linked papers with an AI index.

activity
20242026
collaborators
Showing 2025Show all

2 papers · 1 filter

cs.LG2025

ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm

Rijul Tandon, Peter Vamplew, Cameron Foale

In most value-based reinforcement learning (RL) algorithms, the agent estimates only the expected reward for each action and selects the action with the highest reward. In contrast…

cs.LG2025

On Generalization Across Environments In Multi-Objective Reinforcement Learning

Jayden Teoh, Pradeep Varakantham, Peter Vamplew

Real-world sequential decision-making tasks often require balancing trade-offs between multiple conflicting objectives, making Multi-Objective Reinforcement Learning (MORL) an incr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.