1 paper
Yutong Xin, Qiaochu Chen, Greg Durrett +1
Large language models have achieved striking results in interactive theorem proving, particularly in Lean. However, most benchmarks for LLM-based proof automation are drawn from ma…