1 paper
Chengfeng Dou, Ying Zhang, Zhi Jin +4
This research examines the use of Reinforcement Learning from AI Feedback (RLAIF) techniques to improve healthcare dialogue models, with the aim of tackling the challenges of prefe…