4 papers
A Comparative Study of Neurosymbolic AI Approaches to Interpretable Logical Reasoning
Michael K. Chen
General logical reasoning, defined as the ability to reason deductively on domain-agnostic tasks, continues to be a challenge for large language models (LLMs). Current LLMs fail to…
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI
Huanjin Yao, Jiaxing Huang, Yawen Qiu +9
Reasoning plays a crucial role in advancing Multimodal Large Language Models (MLLMs) toward Artificial General Intelligence. However, existing MLLM benchmarks often fall short in p…
Improving Large Language Models with Concept-Aware Fine-Tuning
Michael K. Chen, Xikun Zhang, Jiaxing Huang +1
Large language models (LLMs) have become the cornerstone of modern AI. However, the existing paradigm of next-token prediction fundamentally limits their ability to form coherent,…
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
Michael K. Chen, Xikun Zhang, Dacheng Tao
Logical reasoning is a critical component of Large Language Models (LLMs), and substantial research efforts in recent years have aimed to enhance their deductive reasoning capabili…