3 papers
cs.CL2025
Can Large Language Models Outperform Non-Experts in Poetry Evaluation? A Comparative Study Using the Consensual Assessment Technique
Piotr Sawicki, Marek GrzeÅ, Dan Brown +1
This study adapts the Consensual Assessment Technique (CAT) for Large Language Models (LLMs), introducing a novel methodology for poetry evaluation. Using a 90-poem dataset with a…
cs.AI2024
Are Frontier Large Language Models Suitable for Q&A in Science Centres?
Jacob Watson, FabrÃcio Góes, Marco Volpe +1
This paper investigates the suitability of frontier Large Language Models (LLMs) for Q&A interactions in science centres, with the aim of boosting visitor engagement while maintain…
cs.AI2024
Do LLMs Agree on the Creativity Evaluation of Alternative Uses?
Abdullah Al Rabeyah, FabrÃcio Góes, Marco Volpe +1
This paper investigates whether large language models (LLMs) show agreement in assessing creativity in responses to the Alternative Uses Test (AUT). While LLMs are increasingly use…