From the 1 of 2 linked papers with an AI index.
1 paper · 1 filter
Zeynel A. UluÅan, Burak S. Akbudak, Can S. Erer +1
Recent neural theorem provers use reinforcement learning with verifiable rewards (RLVR), where proof assistants provide binary correctness signals. While verifiable rewards are che…