3 papers
cs.CY2026
Silent Revision: Measuring Undisclosed Change in the Safety Frameworks of Frontier AI Developers
Louis Yiven Zhu
Frontier AI developers publish safety frameworks that commit them to evidencing whether their models are dangerous. The European Union and California now treat these documents as i…
econ.GN2026
The Price of Intelligence: A Quality-Adjusted Price Index for AI Services
Louis Yiven Zhu
Posted prices for AI inference have fallen steadily since 2024, yet the measured speed of that fall depends almost entirely on the method of measurement. This paper constructs qual…
cs.LG2026
One Capability or Many? Testing the Economic Validity of Frontier AI Evaluation
Louis Yiven Zhu
Frontier-model leaderboards now rank systems based on economic benchmarks, tests of how well models carry out professional tasks from software engineering to banking workflows, and…