2 papers
cs.SI2026
UniShare: A Unified Framework for Joint Video and Receiver Recommendation in Social Sharing
Caimeng Wang, Li Chong, Dongxu Liu +2
Sharing behavior on short-video platforms constitutes a complex ternary interaction among the user (sharer), the video (content), and the receiver. Traditional industrial solutions…
cs.CL2024
PMoL: Parameter Efficient MoE for Preference Mixing of LLM Alignment
Dongxu Liu, Bing Xu, Yinzhuo Chen +4
Reinforcement Learning from Human Feedback (RLHF) has been proven to be an effective method for preference alignment of large language models (LLMs) and is widely used in the post-…