4 papers · 1 filter
Low Rank for Rank: Uncertainty-Aware Task-Specific LLM Ranking under Sparse Pairwise Comparisons
Jiachun Li, David Simchi-Levi, Will Wei Sun
Pairwise human-preference platforms such as Chatbot Arena have become central to large language model (LLM) evaluation, yet reliable task-specific ranking remains challenging. Glob…
LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency
Jiachun Li, David Simchi-Levi, Will Wei Sun
Large language model (LLM) evaluation platforms increasingly rely on pairwise human judgments. These data are noisy, sparse, and non-uniform, yet leaderboards are reported with lim…
Generalized Tensor Completion with Non-Random Missingness
Maoyu Zhang, Biao Cai, Will Wei Sun +1
Tensor completion plays a crucial role in applications such as recommender systems and medical imaging, where data are often highly incomplete. While extensive prior work has addre…
Jointly Modeling and Clustering Tensors in High Dimensions
Biao Cai, Jingfei Zhang, Will Wei Sun
We consider the problem of jointly modeling and clustering populations of tensors by introducing a high-dimensional tensor mixture model with heterogeneous covariances. To effectiv…