Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
CoPRIS: Efficient and Stable Reinforcement Learning via Concurrency-Controlled Partial Rollout with Importance Sampling
Zekai Qu, Yinxu Pan, Ao Sun +2
Reinforcement learning (RL) post-training has become a trending paradigm for enhancing the capabilities of large language models (LLMs). Most existing RL systems for LLMs operate i…
cs.LG2024
Data-centric Graph Learning: A Survey
Yuxin Guo, Deyu Bo, Cheng Yang +5
The history of artificial intelligence (AI) has witnessed the significant impact of high-quality data on various deep learning models, such as ImageNet for AlexNet and ResNet. Rece…