3 papers
cs.AI2025
FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI
Elliot Glazer, Ege Erdil, Tamay Besiroglu +21
We introduce FrontierMath, a benchmark of hundreds of original, exceptionally challenging mathematics problems crafted and vetted by expert mathematicians. The questions cover most…
cs.LG2025
Inference economics of language models
Ege Erdil
We develop a theoretical model that addresses the economic trade-off between cost per token versus serial token generation speed when deploying LLMs for inference at scale. Our mod…
econ.GN2025
GATE: An Integrated Assessment Model for AI Automation
Ege Erdil, Andrei Potlogea, Tamay Besiroglu +6
Assessing the economic impacts of artificial intelligence requires integrating insights from both computer science and economics. We present the Growth and AI Transition Endogenous…