1 paper
Nicole N Khatibi, Daniil A. Radamovich, Michael P. Brenner
Recent breakthroughs have spurred claims that large language models (LLMs) match gold medal Olympiad to graduate level proficiency on mathematics benchmarks. In this work, we exami…