4 papers · 1 filter
The Illusion of Procedural Reasoning: Measuring Long-Horizon FSM Execution in LLMs
Mahdi Samiei, Mahdi Mansouri, Mahdieh Soleymani Baghshah
Large language models (LLMs) have achieved remarkable results on tasks framed as reasoning problems, yet their true ability to perform procedural reasoning, executing multi-step, r…
Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
Mohammad Mahdi Samiei Paqaleh, Arash Marioriyad, Arman Tahmasebi-Zadeh +3
Recent progress has pushed AI frontiers from pattern recognition tasks toward problems that require step by step, System2 style reasoning, especially with large language models. Ye…
MEENA (PersianMMMU): Multimodal-Multilingual Educational Exams for N-level Assessment
Omid Ghahroodi, Arshia Hemmat, Marzia Nouri +8
Recent advancements in large vision-language models (VLMs) have primarily focused on English, with limited attention given to other languages. To address this gap, we introduce MEE…
LLM-Agent-Controller: A Universal Multi-Agent Large Language Model System as a Control Engineer
Rasoul Zahedifar, Sayyed Ali Mirghasemi, Mahdieh Soleymani Baghshah +1
This study presents the LLM-Agent-Controller, a multi-agent large language model (LLM) system developed to address a wide range of problems in control engineering (Control Theory).…