3 papers
cs.LG2026
QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs
Santiago Gonzalez, Alireza Amiri Bavandpour, Peter Ye +48
As Large Language Models (LLMs) saturate elementary benchmarks, the research frontier has shifted from generation to the reliability of automated evaluation. We demonstrate that st…
math.CO2026
Congruent copies of finite patterns in planar point sets
Shubhrajit Bhattacharya, Ritesh Goenka
Given a finite nonempty planar point set , what is the maximum number of congruent copies of contained in a set of points in the Euclidean plane? Building on OpenAI's re…
math.NT2026
Correlations of error terms for weighted prime counting functions
Shubhrajit Bhattacharya, Greg Martin, Reginald M. Simpson
Standard prime-number counting functions, such as , , and , have error terms with limiting logarithmic distributions once suitably normalized. The same is true…