2 citations · 2 across the 1 of their papers we have counts for
1 paper
Sean Williams, James Huckle
We introduce a comprehensive Linguistic Benchmark designed to evaluate the limitations of Large Language Models (LLMs) in domains such as logical reasoning, spatial intelligence, a…