4 papers
Many Voices, One Reward: Multi-Role Rubric Generation for LLM Judging and Reward Modeling
Dazhi Fu, Jiuding Yang, Yiwen Guo +1
Reliable reward and preference signals are critical for evaluating and optimizing large language models on open-ended tasks. Rubric-based judges offer a transparent way to decompos…
Subject Information Extraction for Novelty Detection with Domain Shifts
Yangyang Qu, Dazhi Fu, Jicong Fan
Unsupervised novelty detection (UND), aimed at identifying novel samples, is essential in fields like medical diagnosis, cybersecurity, and industrial quality control. Most existin…
Sample Transform Cost-Based Training-Free Hallucination Detector for Large Language Models
Zeyang Ding, Xinglin Hu, Jicong Fan
Hallucinations in large language models (LLMs) remain a central obstacle to trustworthy deployment, motivating detectors that are accurate, lightweight, and broadly applicable. Sin…
UniOD: A Universal Model for Outlier Detection across Diverse Domains
Dazhi Fu, Jicong Fan
Outlier detection (OD), distinguishing inliers and outliers in completely unlabeled datasets, plays a vital role in science and engineering. Although there have been many insightfu…