Showing 2025Show all
2 papers · 1 filter
cs.CL2025
IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation
Johannes Schmitt, Gergely Bérczi, Jasper Dekoninck +57
As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on research-level tasks at the frontier of…
math.AG2025
A simple criterion for the uniruledness of an orthogonal modular variety
Ignacio Barros
We exhibit a simple uniruledness criterion for general orthogonal modular varieties in terms of invariants of the corresponding lattice. As an application, we obtain the uniruledne…