Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories
Kshitij Gupta
We present a novel data set, WhoDunIt, to assess the deductive reasoning capabilities of large language models (LLM) within narrative contexts. Constructed from open domain mystery…
cs.CL2022
MALM: Mixing Augmented Language Modeling for Zero-Shot Machine Translation
Kshitij Gupta
Large pre-trained language models have brought remarkable progress in NLP. Pre-training and Fine-tuning have given state-of-art performance across tasks in text processing. Data Au…