From the 1 of 18 linked papers with an AI index.
5 citations · 5 across the 6 of their papers we have counts for
6 papers · 1 filter
MathCoPilot: An Interactive System for Human-AI Symbiotic Paradigm of Mathematical Research
Junjie Zhang, Jiayu Liu, Wenbin Liu +11
MathCoPilot is an interactive, human‑in‑the‑loop system that lets mathematicians steer AI agents to formalize and verify mathematical proofs in Lean, combining a live proof bluepri…
UniCog: Uncovering Cognitive Abilities of LLMs through Latent Mind Space Analysis
Jiayu Liu, Yinhe Long, Zhenya Huang +1
A growing body of research suggests that the cognitive processes of large language models (LLMs) differ fundamentally from those of humans. However, existing interpretability metho…
Verifying Large Language Models' Reasoning Paths via Correlation Matrix Rank
Jiayu Liu, Wei Dai, Zhenya Huang +2
Despite the strong reasoning ability of large language models~(LLMs), they are prone to errors and hallucinations. As a result, how to check their outputs effectively and efficient…
Foundation of Intelligence: Review of Math Word Problems from Human Cognition Perspective
Zhenya Huang, Jiayu Liu, Xin Lin +6
Math word problem (MWP) serves as a fundamental research topic in artificial intelligence (AI) dating back to 1960s. This research aims to advance the reasoning abilities of AI by…
CogMath: Assessing LLMs' Authentic Mathematical Ability from a Human Cognitive Perspective
Jiayu Liu, Zhenya Huang, Wei Dai +7
Although large language models (LLMs) show promise in solving complex mathematical tasks, existing evaluation paradigms rely solely on a coarse measure of overall answer accuracy,…
Unveiling the Magic of Code Reasoning through Hypothesis Decomposition and Amendment
Yuze Zhao, Tianyun Ji, Wenjun Feng +6
The reasoning abilities are one of the most enigmatic and captivating aspects of large language models (LLMs). Numerous studies are dedicated to exploring and expanding the boundar…