2 papers
cs.CL2025
ArenaBencher: Automatic Benchmark Evolution via Multi-Model Competitive Evaluation
Qin Liu, Jacob Dineen, Yuxi Huang +4
Benchmarks are central to measuring the capabilities of large language models and guiding model development, yet widespread data leakage from pretraining corpora undermines their v…
math.FA2014
Coarse Quotient Mappings between Metric Spaces
Sheng Zhang
We give a definition of coarse quotient mapping and show that several results for uniform quotient mapping also hold in the coarse setting. In particular, we prove that any Banach…