Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Many Voices, One Reward: Multi-Role Rubric Generation for LLM Judging and Reward Modeling
Dazhi Fu, Jiuding Yang, Yiwen Guo +1
Reliable reward and preference signals are critical for evaluating and optimizing large language models on open-ended tasks. Rubric-based judges offer a transparent way to decompos…
cs.LG2026
CLUBench: A Clustering Benchmark
Feng Xiao, Dazhi Fu, Chris Ding +1
Clustering is a fundamental problem in data science with a long-standing research history, yielding numerous insightful algorithms. Despite this progress, a systematic and large-sc…
cs.LG2025
UniOD: A Universal Model for Outlier Detection across Diverse Domains
Dazhi Fu, Jicong Fan
Outlier detection (OD), distinguishing inliers and outliers in completely unlabeled datasets, plays a vital role in science and engineering. Although there have been many insightfu…