2 papers
cs.CL2026
IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation
Johannes Schmitt, Gergely Bérczi, Jasper Dekoninck +57
As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on research-level tasks at the frontier of…
math.GR2026
Morphisms of generalized affine buildings
Raphael Appenzeller, Xenia Flamm, Victor Jaeck
We define a notion of morphism for generalized affine buildings, also known as affine -buildings, extending existing definitions and giving rise to a category of generalized af…