2 papers
cs.AI2026
PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR
Shumeng Yang, Yisu Liu, Jiayi Zheng +2
Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, where the policy prematurely con…
cs.SD2026
DegDiT: Controllable Audio Generation with Dynamic Event Graph Guided Diffusion Transformer
Yisu Liu, Chenxing Li, Wanqian Zhang +6
Controllable text-to-audio generation aims to synthesize audio from textual descriptions while satisfying user-specified constraints, including event types, temporal sequences, and…