2 papers
cs.AI2025
Performance of LLMs on Stochastic Modeling Operations Research Problems: From Theory to Practice
Akshit Kumar, Tianyi Peng, Yuhang Wu +1
Large language models (LLMs) have exhibited expert-level capabilities across various domains. However, their abilities to solve problems in Operations Research (OR) -- the analysis…
cs.CL2025
Benchmarking Abstract and Reasoning Abilities Through A Theoretical Perspective
Qingchuan Ma, Yuhang Wu, Xiawu Zheng +1
In this paper, we aim to establish a simple, effective, and theoretically grounded benchmark for rigorously probing abstract reasoning in Large Language Models (LLMs). To achieve t…