Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
RubricBench: Aligning Model-Generated Rubrics with Human Standards
Qiyuan Zhang, Junyi Zhou, Yufei Wang +8
As Large Language Model (LLM) alignment evolves from simple completions to complex, highly sophisticated generation, Reward Models are increasingly shifting toward rubric-guided ev…
cs.AI2026
Improving Autoformalization Using Direct Dependency Retrieval
Shaoqi Wang, Lu Yu, Siwei Lou +4
The convergence of deep learning and formal mathematics has spurred research in formal verification. Statement autoformalization, a crucial first step in this process, aims to tran…