Showing 2026Show all
2 papers · 1 filter
cs.CL2026
DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization
Mengyi Deng, Zhiwei Li, Xin Li +4
Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consistency while preserving reason…
cs.IR2026
Quality-Aware Collaborative Multi-Positive Contrastive Learning for Sequential Recommendation
Wei Wang, Yujie Lin, Moyan Zhang +5
The effectiveness of contrastive learning in sequential recommendation hinges on the construction of contrastive views, which ideally should be both semantically consistent and div…