Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
An Interpretable and Scalable Framework for Evaluating Large Language Models
Xinhao Qu, Qiang Heng, Hao Zeng +1
Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the inherent stochasticity of LLM…
stat.ML2025
Inertial Quadratic Majorization Minimization with Application to Kernel Regularized Learning
Qiang Heng, Caixing Wang
First-order methods in convex optimization offer low per-iteration cost but often suffer from slow convergence, while second-order methods achieve fast local convergence at the exp…