3 papers
cs.CL2024
Assessing Language Models' Worldview for Fiction Generation
Aisha Khatun, Daniel G. Brown
The use of Large Language Models (LLMs) has become ubiquitous, with abundant applications in computational creativity. One such application is fictional story generation. Fiction i…
cs.CL2024
A Study on Large Language Models' Limitations in Multiple-Choice Question Answering
Aisha Khatun, Daniel G. Brown
The widespread adoption of Large Language Models (LLMs) has become commonplace, particularly with the emergence of open-source models. More importantly, smaller models are well-sui…
cs.CL2024
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
Aisha Khatun, Daniel G. Brown
Large Language Model (LLM) evaluation is currently one of the most important areas of research, with existing benchmarks proving to be insufficient and not completely representativ…