2 papers
cs.CL2024
Reinforcement Learning with Token-level Feedback for Controllable Text Generation
Wendi Li, Wei Wei, Kaihe Xu +3
To meet the requirements of real-world applications, it is essential to control generations of large language models (LLMs). Prior research has tried to introduce reinforcement lea…
cs.IR2023
Multi-view Hypergraph Contrastive Policy Learning for Conversational Recommendation
Sen Zhao, Wei Wei, Xian-Ling Mao +5
Conversational recommendation systems (CRS) aim to interactively acquire user preferences and accordingly recommend items to users. Accurately learning the dynamic user preferences…