2 papers
cs.LG2026
Elucidating Representation Degradation Problem in Diffusion Model Training
Zhipeng Yao, Dazhou Li, Zitong Zhang +6
Diffusion models have achieved remarkable success, yet their training remains inefficient due to a severe optimization bottleneck, which we term Representation Degradation. As nois…
cs.CL2025
AIR: A Systematic Analysis of Annotations, Instructions, and Response Pairs in Preference Dataset
Bingxiang He, Wenbin Zhang, Jiaxi Song +11
Preference learning is critical for aligning large language models (LLMs) with human values, yet its success hinges on high-quality datasets comprising three core components: Prefe…