4 papers
Many Voices, One Reward: Multi-Role Rubric Generation for LLM Judging and Reward Modeling
Dazhi Fu, Jiuding Yang, Yiwen Guo +1
Reliable reward and preference signals are critical for evaluating and optimizing large language models on open-ended tasks. Rubric-based judges offer a transparent way to decompos…
CLUBench: A Clustering Benchmark
Feng Xiao, Dazhi Fu, Chris Ding +1
Clustering is a fundamental problem in data science with a long-standing research history, yielding numerous insightful algorithms. Despite this progress, a systematic and large-sc…
Subject Information Extraction for Novelty Detection with Domain Shifts
Yangyang Qu, Dazhi Fu, Jicong Fan
Unsupervised novelty detection (UND), aimed at identifying novel samples, is essential in fields like medical diagnosis, cybersecurity, and industrial quality control. Most existin…
UniOD: A Universal Model for Outlier Detection across Diverse Domains
Dazhi Fu, Jicong Fan
Outlier detection (OD), distinguishing inliers and outliers in completely unlabeled datasets, plays a vital role in science and engineering. Although there have been many insightfu…