13 papers
APEX-VW: A Document-Level English-Spanish Post-Editing Dataset in the Healthcare Domain
Marie Escribe, Tharindu Ranasinghe, Amal Haddad Haddad +2
Post-Editing (PE) of Machine Translation (MT) output often involves repeating the same lexical and terminological corrections across many segments, especially in specialised and hi…
ltzGLUE: Luxembourgish General Language Understanding Evaluation
Alistair Plum, Felicia Körner, Anne-Marie Lutgen +8
This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for English. Although NLU tasks ar…
Do LLMs Judge Distantly Supervised Named Entity Labels Well? Constructing the JudgeWEL Dataset
Alistair Plum, Laura Bernardy, Tharindu Ranasinghe
We present judgeWEL, a dataset for named entity recognition (NER) in Luxembourgish, automatically labelled and subsequently verified using large language models (LLM) in a novel pi…
MUNIChus: Multilingual News Image Captioning Benchmark
Yuji Chen, Alistair Plum, Hansi Hettiarachchi +4
The goal of news image captioning is to generate captions by integrating news article content with corresponding images, highlighting the relationship between textual context and v…
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
Tharindu Cyril Weerasooriya, Sujan Dutta, Tharindu Ranasinghe +3
Offensive speech detection is a key component of content moderation. However, what is offensive can be highly subjective. This paper investigates how machine and human moderators d…
A Survey on Multilingual Mental Disorders Detection from Social Media Data
Ana-Maria Bucur, Marcos Zampieri, Tharindu Ranasinghe +1
The increasing prevalence of mental disorders globally highlights the urgent need for effective digital screening methods that can be used in multilingual contexts. Most existing s…