Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Winning Gold at IMO 2025 with a Model-Agnostic Verification-and-Refinement Pipeline
Yichen Huang, Lin F. Yang
The International Mathematical Olympiad (IMO) is widely regarded as the world championship of high-school mathematics. IMO problems are renowned for their difficulty and novelty, d…
cs.AI2025
Do LLMs Know When to Flip a Coin? Strategic Randomization through Reasoning and Experience
Lingyu Yang
Strategic randomization is a key principle in game theory, yet it remains underexplored in large language models (LLMs). Prior work often conflates the cognitive decision to random…