3 papers
cs.AI2025
Verifying Large Language Models' Reasoning Paths via Correlation Matrix Rank
Jiayu Liu, Wei Dai, Zhenya Huang +2
Despite the strong reasoning ability of large language models~(LLMs), they are prone to errors and hallucinations. As a result, how to check their outputs effectively and efficient…
cs.AI2025
CogMath: Assessing LLMs' Authentic Mathematical Ability from a Human Cognitive Perspective
Jiayu Liu, Zhenya Huang, Wei Dai +7
Although large language models (LLMs) show promise in solving complex mathematical tasks, existing evaluation paradigms rely solely on a coarse measure of overall answer accuracy,…
cs.CR2025
Valida ISA Spec, version 1.0: A zk-Optimized Instruction Set Architecture
Morgan Thomas, Mamy Ratsimbazafy, Marcin Bugaj +7
The Valida instruction set architecture is designed for implementation in zkVMs to optimize for fast, efficient execution proving. This specification intends to guide implementors…