23 citations · 24 across the 4 of their papers we have counts for
4 papers · 1 filter
Reward-Driven Interaction: Enhancing Proactive Dialogue Agents through User Satisfaction Prediction
Wei Shen, Xiaonan He, Chuheng Zhang +3
Reward-driven proactive dialogue agents require precise estimation of user satisfaction as an intrinsic reward signal to determine optimal interaction strategies. Specifically, thi…
Policy Filtration for RLHF to Mitigate Noise in Reward Models
Chuheng Zhang, Wei Shen, Li Zhao +4
While direct policy optimization methods exist, pioneering LLMs are fine-tuned with reinforcement learning from human feedback (RLHF) to generate better responses under the supervi…
Inductive Matrix Completion Using Graph Autoencoder
Wei Shen, Chuheng Zhang, Yun Tian +4
Recently, the graph neural network (GNN) has shown great power in matrix completion by formulating a rating matrix as a bipartite graph and then predicting the link between the cor…
Auxiliary-task Based Deep Reinforcement Learning for Participant Selection Problem in Mobile Crowdsourcing
Wei Shen, Xiaonan He, Chuheng Zhang +3
In mobile crowdsourcing (MCS), the platform selects participants to complete location-aware tasks from the recruiters aiming to achieve multiple goals (e.g., profit maximization, e…