large language model evaluation 1preference learning 1query-only supervision 1rubric generation 1synthetic pairwise data 1
From the 1 of 6 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Rubrics on Trial: Evolving Rubrics from a Single Query via Synthetic Pairwise Evidence
Haocheng Yang, Licheng Pan, Xiaoxi Li +5
The paper proposes a query‑only method that automatically creates and validates fine‑grained rubrics for evaluating large language models by using synthetic rubric‑conditioned resp…
cs.CL2025
Mitigating Hidden Confounding by Progressive Confounder Imputation via Large Language Models
Hao Yang, Haoxuan Li, Luyu Chen +3
Hidden confounding remains a central challenge in estimating treatment effects from observational data, as unobserved variables can lead to biased causal estimates. While recent wo…