activity
20242026
collaborators

9 papers

cs.CL2026

How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP

Kushal Tatariya, Artur Kulmizev, Wessel Poelman +6

Wikipedia's perceived high quality and broad language coverage have established it as a fundamental resource in NLP. However, in recent years, such assumptions of high quality have…

cs.CL2026

Form and Meaning in Intrinsic Multilingual Evaluations

Wessel Poelman, Miryam de Lhoneux

Intrinsic evaluation metrics for conditional language models, such as perplexity or bits-per-character, are widely used in both mono- and multilingual settings. These metrics are r…

cs.CL2025

On the Interplay between Positional Encodings, Morphological Complexity, and Word Order Flexibility

Kushal Tatariya, Wessel Poelman, Miryam de Lhoneux

Language model architectures are predominantly first created for English and subsequently applied to other languages. It is an open question whether this architectural bias leads t…

cs.CL2025

Confounding Factors in Relating Model Performance to Morphology

Wessel Poelman, Thomas Bauwens, Miryam de Lhoneux

The extent to which individual language characteristics influence tokenization and language modeling is an open question. Differences in morphological systems have been suggested a…

cs.CL2025

A Principled Framework for Evaluating on Typologically Diverse Languages

Esther Ploeger, Wessel Poelman, Andreas Holck Høeg-Petersen +3

Beyond individual languages, multilingual natural language processing (NLP) research increasingly aims to develop models that perform well across languages generally. However, eval…

cs.CL2024

The Roles of English in Evaluating Multilingual Language Models

Wessel Poelman, Miryam de Lhoneux

Multilingual natural language processing is getting increased attention, with numerous models, benchmarks, and methods being released for many languages. English is often used in m…