13 citations · 16 across the 2 of their papers we have counts for
2 papers
cs.LG2024★ 3 cited
HRLAIF: Improvements in Helpfulness and Harmlessness in Open-domain Reinforcement Learning From AI Feedback
Ang Li, Qiugen Xiao, Peng Cao +12
Reinforcement Learning from AI Feedback (RLAIF) has the advantages of shorter annotation cycles and lower costs over Reinforcement Learning from Human Feedback (RLHF), making it hi…
gr-qc2022★ 13 cited
Generation of quantum coherence for continuous variables between causally disconnected regions in dilaton spacetime
Qinglong Xiao, Cuihong Wen, Jiliang Jing +1
We study the dynamics of Gaussian quantum coherence under the background of a Garfinkle-Horowitz-Strominger dilaton black hole. It is shown that the dilaton field has evident effec…