1 paper · 1 filter
Fien van Wetten, Aske Plaat, Max van Duijn
Large language models (LLMs) are known to perform well on language tasks, but struggle with reasoning tasks. This paper explores the ability of LLMs to play the 2D puzzle game Baba…