Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
From Atomic to Agentic: Towards Interpretable Evaluation of LLMs' Agentic Mathematical Capabilities
Jiayi Kuang, Yinghui Li, Yunze Song +11
Large Language Models (LLMs) are evolving from performing end-to-end mathematical reasoning to integrating agentic intelligence. However, most existing math benchmarks evaluate onl…
cs.AI2025
Adaptive Dual Reasoner: Large Reasoning Models Can Think Efficiently by Hybrid Reasoning
Yujian Zhang, Keyu Chen, Zhifeng Shen +2
Although Long Reasoning Models (LRMs) have achieved superior performance on various reasoning scenarios, they often suffer from increased computational costs and inference latency…