4 papers · 1 filter
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
Soniya Vijayakumar, Josef van Genabith, Simon Ostermann
In the era of high performing Large Language Models, researchers have widely acknowledged that contextual word representations are one of the key drivers in achieving top performan…
MCScript2.0: A Machine Comprehension Corpus Focused on Script Events and Participants
Simon Ostermann, Michael Roth, Manfred Pinkal
We introduce MCScript2.0, a machine comprehension corpus for the end-to-end evaluation of script knowledge. MCScript2.0 contains approx. 20,000 questions on approx. 3,500 texts, cr…
MCScript: A Novel Dataset for Assessing Machine Comprehension Using Script Knowledge
Simon Ostermann, Ashutosh Modi, Michael Roth +2
We introduce a large dataset of narrative texts and questions about these texts, intended to be used in a machine comprehension task that requires reasoning using commonsense knowl…
InScript: Narrative texts annotated with script information
Ashutosh Modi, Tatjana Anikina, Simon Ostermann +1
This paper presents the InScript corpus (Narrative Texts Instantiating Script structure). InScript is a corpus of 1,000 stories centered around 10 different scenarios. Verbs and no…