2 papers
cs.CL2026
Self-Verification is All You Need To Pass The Japanese Bar Examination
Andrew Shin
Despite rapid advances in large language models (LLMs), achieving reliable performance on highly professional and structured examinations remains a significant challenge. The Japan…
cs.CL2025
Can A Gamer Train A Mathematical Reasoning Model?
Andrew Shin
While large language models (LLMs) have achieved remarkable performance in various tasks including mathematical reasoning, their development typically demands prohibitive computati…