3 papers
cs.AI2025
The Procedural Content Generation Benchmark: An Open-source Testbed for Generative Challenges in Games
Ahmed Khalifa, Roberto Gallotta, Matthew Barthet +3
This paper introduces the Procedural Content Generation Benchmark for evaluating generative algorithms on different game content creation tasks. The benchmark comes with 12 game-re…
cs.SE2025
FREYR: A Framework for Recognizing and Executing Your Requests
Roberto Gallotta, Antonios Liapis, Georgios N. Yannakakis
Large language models excel as conversational agents, but their capabilities can be further extended through tool usage, i.e.: executable code, to enhance response accuracy or addr…
cs.CL2024
Large Language Models and Games: A Survey and Roadmap
Roberto Gallotta, Graham Todd, Marvin Zammit +4
Recent years have seen an explosive increase in research on large language models (LLMs), and accompanying public engagement on the topic. While starting as a niche area within nat…