2 papers
cs.CL2026
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes
Avinash Anand, Mahisha Ramesh, Avni Mittal +8
Reasoning has become central to how Large Language Models (LLMs) are evaluated and interpreted, spanning Chain-of-Thought (CoT), mathematical problem-solving, multi-hop question an…
cs.MA2025
Beyond Task Completion: An Assessment Framework for Evaluating Agentic AI Systems
Sreemaee Akshathala, Bassam Adnan, Mahisha Ramesh +3
Recent advances in agentic AI have shifted the focus from standalone Large Language Models (LLMs) to integrated systems that combine LLMs with tools, memory, and other agents to pe…