◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Benjamin Van Roy

Stanford University

4 papers hereh-index 5617.1k citations204 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author3

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG4
affiliations
  • Stanford University
Homepage
same name
  • Benjamin Van Roy — 4 papers, h 3
  • Benjamin Van Roy — 4 papers, h 3
  • Benjamin Van Roy — 3 papers, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedNon-Stationary Bandit Learning via Predictive Sampling

6 citations · 6 across the 1 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026★ 6 cited

Non-Stationary Bandit Learning via Predictive Sampling

Yueyang Liu, Xu Kuang, Benjamin Van Roy

Thompson sampling has proven effective across a wide range of stationary bandit environments. However, as we demonstrate in this paper, it can perform poorly when applied to non-st…

cs.LG2025

Posterior Sampling for Continuing Environments

Wanqiao Xu, Shi Dong, Benjamin Van Roy

We develop an extension of posterior sampling for reinforcement learning (PSRL) that is suited for a continuing agent-environment interface and integrates naturally into agent desi…

cs.LG2025

Continual Learning as Computationally Constrained Reinforcement Learning

Saurabh Kumar, Henrik Marklund, Ashish Rao +4

An agent that efficiently accumulates knowledge to develop increasingly sophisticated skills over a long lifetime could advance the frontier of artificial intelligence capabilities…

cs.LG2024

Maintaining Plasticity in Continual Learning via Regenerative Regularization

Saurabh Kumar, Henrik Marklund, Benjamin Van Roy

In continual learning, plasticity refers to the ability of an agent to quickly adapt to new information. Neural networks are known to lose plasticity when processing non-stationary…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.