2 papers
cs.AI2025
Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution
Zailong Tian, Zhuoheng Han, Yanzhe Chen +5
Large Language Models (LLMs) are widely used as automated judges, where practical value depends on both accuracy and trustworthy, risk-aware judgments. Existing approaches predomin…
cs.IR2025
Beyond Single Labels: Improving Conversational Recommendation through LLM-Powered Data Augmentation
Haozhe Xu, Xiaohua Wang, Changze Lv +1
Conversational recommender systems (CRSs) enhance recommendation quality by engaging users in multi-turn dialogues, capturing nuanced preferences through natural language interacti…