2 papers
cs.LG2024
Understanding and Alleviating Memory Consumption in RLHF for LLMs
Jin Zhou, Hanmei Yang, Steven +4
Fine-tuning with Reinforcement Learning with Human Feedback (RLHF) is essential for aligning large language models (LLMs). However, RLHF often encounters significant memory challen…
cs.PF2024
Scaler: Efficient and Effective Cross Flow Analysis
Steven, Tang, Mingcan Xiang +4
Performance analysis is challenging as different components (e.g.,different libraries, and applications) of a complex system can interact with each other. However, few existing too…