45 citations · 45 across the 2 of their papers we have counts for
2 papers
cs.LG2025
Learning a Canonical Basis of Human Preferences from Binary Ratings
Kailas Vodrahalli, Wei Wei, James Zou
Recent advances in generative AI have been driven by alignment techniques such as reinforcement learning from human feedback (RLHF). RLHF and related techniques typically involve c…
cs.LG2023★ 45 cited
Can large language models provide useful feedback on research papers? A large-scale empirical analysis
Weixin Liang, Yuhui Zhang, Hancheng Cao +9
Expert feedback lays the foundation of rigorous research. However, the rapid growth of scholarly production and intricate knowledge specialization challenge the conventional scient…