1 paper
Qingchuan Ma, Yuhang Wu, Xiawu Zheng +1
In this paper, we aim to establish a simple, effective, and theoretically grounded benchmark for rigorously probing abstract reasoning in Large Language Models (LLMs). To achieve t…