Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
DiffPO: Diffusion-styled Preference Optimization for Efficient Inference-Time Alignment of Large Language Models
Ruizhe Chen, Wenhao Chai, Zhifei Yang +5
Inference-time alignment provides an efficient alternative for aligning LLMs with humans. However, these approaches still face challenges, such as limited scalability due to policy…
cs.CL2024
PAD: Personalized Alignment of LLMs at Decoding-Time
Ruizhe Chen, Xiaotian Zhang, Meng Luo +2
Aligning with personalized preferences, which vary significantly across cultural, educational, and political differences, poses a significant challenge due to the computational cos…