2 papers
cs.AI2026
Reliability-Aware LLM Alignment from Inconsistent Human Feedback
Jingyi Huang, Ruohan Zong, Yujun Feng +3
Reinforcement Learning from Human Feedback (RLHF) is critical for aligning Large Language Models (LLMs) with human preferences. However, its efficacy is often compromised by the in…
cs.IR2025
Use of Retrieval-Augmented Large Language Model Agent for Long-Form COVID-19 Fact-Checking
Jingyi Huang, Yuyi Yang, Mengmeng Ji +3
The COVID-19 infodemic calls for scalable fact-checking solutions that handle long-form misinformation with accuracy and reliability. This study presents SAFE (system for accurate…