10 citations · 10 across the 1 of their papers we have counts for
1 paper · 1 filter
Yanrui Wu, Lingling Zhang, Xinyu Zhang +5
Evaluations of large language models (LLMs) primarily emphasize convergent logical reasoning, where success is defined by producing a single correct proof. However, many real-world…