◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tang

4 papers

No researched profile yet.

papers

Publications (4)

cs.LG2024

Understanding and Alleviating Memory Consumption in RLHF for LLMs

Jin Zhou, Hanmei Yang, Steven +4

Fine-tuning with Reinforcement Learning with Human Feedback (RLHF) is essential for aligning large language models (LLMs). However, RLHF often encounters significant memory challen…

cs.AI2026

CatalogAgent: A Supervisor-mediated Self-Learning System Enabling Context Engineering for GenAI Models

Zhu Cheng, Zhenming Wang, Yu +14

The paper presents CatalogAgent, an agentic system that uses a Supervisor Agent to resolve conflicts between LLM-based generators and evaluators for filling missing product attribu…

#e-commerce#product catalog enrichment#large language models#agentic systems
cs.PF2024

Scaler: Efficient and Effective Cross Flow Analysis

Steven, Tang, Mingcan Xiang +4

Performance analysis is challenging as different components (e.g.,different libraries, and applications) of a complex system can interact with each other. However, few existing too…

cs.PF2022

CachePerf: A Unified Cache Miss Classifier via Hybrid Hardware Sampling

Jin Zhou, Steven, Tang +2

The cache plays a key role in determining the performance of applications, no matter for sequential or concurrent programs on homogeneous and heterogeneous architecture. Fixing cac…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Sign in
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.