2 papers
cs.LG2025
Refining Alignment Framework for Diffusion Models with Intermediate-Step Preference Ranking
Jie Ren, Yuhang Zhang, Dongrui Liu +2
Direct preference optimization (DPO) has shown success in aligning diffusion models with human preference. Previous approaches typically assume a consistent preference label betwee…
cs.CL2024
Identifying Semantic Induction Heads to Understand In-Context Learning
Jie Ren, Qipeng Guo, Hang Yan +4
Although large language models (LLMs) have demonstrated remarkable performance, the lack of transparency in their inference logic raises concerns about their trustworthiness. To ga…