3 papers
cs.AI2026
Reliability-Aware LLM Alignment from Inconsistent Human Feedback
Jingyi Huang, Ruohan Zong, Yujun Feng +3
Reinforcement Learning from Human Feedback (RLHF) is critical for aligning Large Language Models (LLMs) with human preferences. However, its efficacy is often compromised by the in…
cs.CL2025
GAP: Graph-Assisted Prompts for Dialogue-based Medication Recommendation
Jialun Zhong, Yanzeng Li, Sen Hu +3
Medication recommendations have become an important task in the healthcare domain, especially in measuring the accuracy and safety of medical dialogue systems (MDS). Different from…
cs.CL2025
A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future
Jialun Zhong, Wei Shen, Yanzeng Li +7
Reward Model (RM) has demonstrated impressive potential for enhancing Large Language Models (LLM), as RM can serve as a proxy for human preferences, providing signals to guide LLMs…