1 paper
Jiaze Li, Aocheng Shen, Bing Liu +4
Competitive programming is increasingly being used to evaluate the algorithmic reasoning capabilities of large language models (LLMs). However, existing benchmarks primarily focus…