activity
20242026
collaborators
Showing cs.CLShow all

12 papers · 1 filter

cs.CL2026

A Study of Crosslinguistic Influence in Language Models

Abderrahmane Issam, Yusuf Can Semerci, Jan Scholtes +1

The sequential acquisition of languages inevitably leads to Crosslinguistic Influence (CLI), where the syntactic properties of a first language (L1) impact the processing of a seco…

cs.CL2025

Evaluating LLM-Generated Legal Explanations for Regulatory Compliance in Social Media Influencer Marketing

Haoyang Gui, Thales Bertaglia, Taylor Annabell +3

The rise of influencer marketing has blurred boundaries between organic content and sponsored content, making the enforcement of legal rules relating to transparency challenging. E…

cs.CL2025

DTW-Align: Bridging the Modality Gap in End-to-End Speech Translation with Dynamic Time Warping Alignment

Abderrahmane Issam, Yusuf Can Semerci, Jan Scholtes +1

End-to-End Speech Translation (E2E-ST) is the task of translating source speech directly into target text bypassing the intermediate transcription step. The representation discrepa…

cs.CL2025

You Are What You Train: Effects of Data Composition on Training Context-aware Machine Translation Models

Paweł MÄ ka, Yusuf Can Semerci, Jan Scholtes +1

Achieving human-level translations requires leveraging context to ensure coherence and handle complex phenomena like pronoun disambiguation. Sparsity of contextually rich examples…

cs.CL2025

Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch

Elza Strazda, Gerasimos Spanakis

Warning: This paper contains explicit statements of offensive stereotypes which might be upsetting. Language models are prone to exhibiting biases, further amplifying unfair and ha…

cs.CL2025

A Representation Level Analysis of NMT Model Robustness to Grammatical Errors

Abderrahmane Issam, Yusuf Can Semerci, Jan Scholtes +1

Understanding robustness is essential for building reliable NLP systems. Unfortunately, in the context of machine translation, previous work mainly focused on documenting robustnes…