2 papers
stat.ML2026
Low-Rank Contextual Reinforcement Learning from Heterogeneous Human Feedback
Seong Jin Lee, Will Wei Sun, Yufeng Liu
Reinforcement learning from human feedback (RLHF) has become a cornerstone for aligning large language models with human preferences. However, the heterogeneity of human feedback,…
cs.IR2026
Low-Rank Online Dynamic Assortment with Dual Contextual Information
Seong Jin Lee, Will Wei Sun, Yufeng Liu
As e-commerce expands, delivering real-time personalized recommendations from vast catalogs poses a critical challenge for retail platforms. Maximizing revenue requires careful con…