Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Process-Verified Reinforcement Learning for Theorem Proving via Lean
Minsu Kim, Se-Young Yun
While reinforcement learning from verifiable rewards (RLVR) typically has relied on a single binary verification signal, symbolic proof assistants in formal reasoning offer rich, f…
cs.AI2026
Generative Recursive Reasoning
Junyeob Baek, Mingyu Jo, Minsu Kim +3
How should future neural reasoning systems implement extended computation? Recursive Reasoning Models (RRMs) offer a promising alternative to autoregressive sequence extension by p…