3 papers
econ.GN2026
On Benchmark Hacking in ML Contests: Modeling, Insights and Design
Xiaoyun Qiu, Yang Yu, Haifeng Xu
Benchmark hacking refers to tuning a machine learning model to score highly on certain evaluation criteria without improving true generalization or faithfully solving the intended…
cs.GT2026
The Complexity of Tullock Contests
Yu He, Fan Yao, Yang Yu +3
Despite the extensive literature on Tullock contests, computational results for the general model with heterogeneous contestants remain scarce. This paper studies the algorithmic c…
econ.TH2025
Multi-dimensional Test Design
Xiaoyun Qiu, Liren Shan
How should one jointly design tests and the arrangement of agencies to administer these tests (testing procedure)? To answer this question, we analyze a model where a principal mus…