◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Andy Peng

4 papers hereh-index 2102 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.CV1

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedEfficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing 2026 · cs.LGShow all

2 papers · 2 filters

cs.LG2026

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning

Zhiyuan Zhou, Andy Peng, Charles Xu +4

Expressive continuous control policies, such as diffusion and flow models, form the backbone of recent advances in scaling imitation learning for simulated and real robot control.…

cs.LG2026

Fisher-Orthogonal Projected Natural Gradient Descent for Continual Learning

Ishir Garg, Neel Kolhe, Andy Peng +1

Continual learning aims to enable neural networks to acquire new knowledge on sequential tasks. However, the key challenge in such settings is to learn new tasks without catastroph…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.