◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Peter Stone

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.RO2
same name
  • Peter Stone — 22 papers
  • Peter Stone — 5 papers
  • Peter Stone — 4 papers
  • Peter Stone — 4 papers, h 6
  • Peter Stone — 4 papers
  • Peter Stone — 3 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedBenchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks

1 citations · 1 across the 4 of their papers we have counts for

collaborators

4 papers

cs.RO2025

L3M+P: Lifelong Planning with Large Language Models

Krish Agarwal, Yuqian Jiang, Jiaheng Hu +2

By combining classical planning methods with large language models (LLMs), recent research such as LLM+P has enabled agents to plan for general tasks given in natural language. How…

cs.RO2025★ 1 cited

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks

Viraj Joshi, Zifan Xu, Bo Liu +2

Multi-task Reinforcement Learning (MTRL) has emerged as a critical training paradigm for applying reinforcement learning (RL) to a set of complex real-world robotic tasks, which de…

cs.LG2024

Learning Memory Mechanisms for Decision Making through Demonstrations

William Yue, Bo Liu, Peter Stone

In Partially Observable Markov Decision Processes, integrating an agent's history into memory poses a significant challenge for decision-making. Traditional imitation learning, rel…

cs.LG2024

Fine-Grained Gradient Restriction: A Simple Approach for Mitigating Catastrophic Forgetting

Bo Liu, Mao Ye, Peter Stone +1

A fundamental challenge in continual learning is to balance the trade-off between learning new tasks and remembering the previously acquired knowledge. Gradient Episodic Memory (GE…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.