3 papers
cs.LG2026
QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs
Santiago Gonzalez, Alireza Amiri Bavandpour, Peter Ye +48
As Large Language Models (LLMs) saturate elementary benchmarks, the research frontier has shifted from generation to the reliability of automated evaluation. We demonstrate that st…
math.NT2026
Arithmetic exceptionality of Lattès maps
Chatchawan Panraksa, Detchat Samart, Songpon Sriwongsa
Let denote a finite field of order . A rational function is said to be arithmetically exceptional if it induces a permutation on $\mathbb{…
math.NT2025
Determinants of Mahler measures and special values of -functions
Detchat Samart, Zhengyu Tao
We consider Mahler measures of two well-studied families of bivariate polynomials, namely and , where is a complex…