Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
DARTS: Distribution-Aware Active Rollout Trajectory Shaping for Accelerating LLM Reinforcement Learning
Yujie Wang, Siwei Chen, Longzan Luo +4
Reinforcement Learning (RL) has become pivotal for improving model capabilities yet suffers from rollout efficiency bottlenecks due to the long-tail response length distribution. W…
cs.LG2025
Do Graph Diffusion Models Accurately Capture and Generate Substructure Distributions?
Xiyuan Wang, Yewei Liu, Lexi Pang +2
Diffusion models have gained popularity in graph generation tasks; however, the extent of their expressivity concerning the graph distributions they can learn is not fully understo…