activity
20182022
most citedKeyphrase Generation for Scientific Document Retrieval

34 citations · 86 across the 19 of their papers we have counts for

collaborators

30 papers

cs.IR2023

TEIMMA: The First Content Reuse Annotator for Text, Images, and Math

Ankit Satpute, André Greiner-Petter, Moritz Schubotz +4

This demo paper presents the first tool to annotate the reuse of text, images, and mathematical formulae in a document pair -- TEIMMA. Annotating content reuse is particularly usef…

cs.CL2022

Which Shortcut Solution Do Question Answering Models Prefer to Learn?

Kazutoshi Shinoda, Saku Sugawara, Akiko Aizawa

Question answering (QA) models for reading comprehension tend to learn shortcut solutions rather than the solutions intended by QA datasets. QA models that have learned shortcut so…

cs.CL2022

Penalizing Confident Predictions on Largely Perturbed Inputs Does Not Improve Out-of-Distribution Generalization in Question Answering

Kazutoshi Shinoda, Saku Sugawara, Akiko Aizawa

Question answering (QA) models are shown to be insensitive to large perturbations to inputs; that is, they make correct and confident predictions even when given largely perturbed…

cs.SE20222 cited

Caching and Reproducibility: Making Data Science experiments faster and FAIRer

Moritz Schubotz, Ankit Satpute, Andre Greiner-Petter +2

Small to medium-scale data science experiments often rely on research software developed ad-hoc by individual scientists or small teams. Often there is no time to make the research…

cs.CL202225 cited

Exploiting Transformer-based Multitask Learning for the Detection of Media Bias in News Articles

Timo Spinde, Jan-David Krieger, Terry Ruas +4

Media has a substantial impact on the public perception of events. A one-sided or polarizing perspective on any topic is usually described as media bias. One of the ways how bias i…

cs.CL20222 cited

Debiasing Masks: A New Framework for Shortcut Mitigation in NLU

Johannes Mario Meissner, Saku Sugawara, Akiko Aizawa

Debiasing language models from unwanted behaviors in Natural Language Understanding tasks is a topic with rapidly increasing interest in the NLP community. Spurious statistical cor…