◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

S. Tang

3 papers hereh-index 00 citations3 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1

Across the 1 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.PF1
same name
  • S. Tang — 2 papers, h 1
  • S. Tang — 1 paper, h 12
  • S. Tang — 1 paper, h 0
  • S. Tang — 1 paper
  • S. Tang — 1 paper, h 1
  • S. Tang — 1 paper, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.LG2026

Forager: a lightweight testbed for continual learning with partial observability in RL

Steven Tang, Xinze Xiong, Anna Hakhverdyan +7

In continual reinforcement learning (CRL), good performance requires never-ending learning, acting, and exploration in a big, partially observable world. Most CRL experiments have…

cs.LG2024

Understanding and Alleviating Memory Consumption in RLHF for LLMs

Jin Zhou, Hanmei Yang, Steven +4

Fine-tuning with Reinforcement Learning with Human Feedback (RLHF) is essential for aligning large language models (LLMs). However, RLHF often encounters significant memory challen…

cs.PF2024

Scaler: Efficient and Effective Cross Flow Analysis

Steven, Tang, Mingcan Xiang +4

Performance analysis is challenging as different components (e.g.,different libraries, and applications) of a complex system can interact with each other. However, few existing too…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.