1 paper
Yuchen Li, Cong Lin, Muhammad Umair Nasir +3
We introduce GVGAI-LLM, a video game benchmark for evaluating the reasoning and problem-solving capabilities of large language models (LLMs). Built on the General Video Game AI fra…