3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2024★ 3 cited
HRLAIF: Improvements in Helpfulness and Harmlessness in Open-domain Reinforcement Learning From AI Feedback
Ang Li, Qiugen Xiao, Peng Cao +12
Reinforcement Learning from AI Feedback (RLAIF) has the advantages of shorter annotation cycles and lower costs over Reinforcement Learning from Human Feedback (RLHF), making it hi…
cs.IR2023
Binary Embedding-based Retrieval at Tencent
Yukang Gan, Yixiao Ge, Chang Zhou +7
Large-scale embedding-based retrieval (EBR) is the cornerstone of search-related industrial applications. Given a user query, the system of EBR aims to identify relevant informatio…