3 papers
cs.AI2026
Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates
Bohan Chen, Shivam N. Patel, Richard Hoffmann +2
Tool calling allows large language models (LLMs) to invoke external computation during problem solving, a useful capability in various fields including AI for mathematics. We study…
cs.AI2026
An AI Scientist that Doesn't Drift: Taste, Structure, and Falsifiable Findings in a Quadruped Navigation Research Loop
Yiwen Zhang, Eloise Zeng, Jaeha Lee +1
Autonomous research loops driven by large language models can run machine-learning experiments at scale but tend to drift toward local refinements of whichever metric they optimise…
cs.LG2026
AlphaZero in Sparsely Rewarded Games: Limits and Auxiliary Supervision
Brent Kong, Tejas Ram, Tony Yue Yu
AlphaZero has demonstrated that a neural-guided Monte Carlo Tree Search can achieve superhuman performance, but strong play does not necessarily imply perfect play. We study this g…