2 papers
cs.CL2026
Can LLM Reasoning Be Trusted? A Comparative Study: Using Human Benchmarking on Statistical Tasks
Crish Nagarkar, Leonid Bogachev, Serge Sharoff
This paper investigates the ability of large language models (LLMs) to solve statistical tasks, as well as their capacity to assess the quality of reasoning. While state-of-the-art…
cs.DL2026
A citation index bridging Hirsch's h and Egghe's g
Ruheyan Nuermaimaiti, Leonid V. Bogachev, Jochen Voss
We propose a citation index (``nu'') and show that it lies between the classical -index and -index. This idea is then generalized to a monotone parametric family $(ν_α…