collaborators

8 papers

cs.CL2026

Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining

Felicia Körner, Maria Matveev, Florian Eichin +3

Large language models exhibit impressive cross-lingual capabilities. However, prior work analyzes this phenomenon through isolated factors and at sparse points during training, lim…

cs.CV2026

Edit the Bits, Diff the Codes: Bitwise Residual Editing for Visual Autoregressive Models

Shengqiang Zhang, Ruotong Liao, Volker Tresp +2

Text-guided image editing with visual autoregressive (VAR) generators requires controlling both what the model samples and where the sampled change is written back into the image c…

cs.CL2026

ltzGLUE: Luxembourgish General Language Understanding Evaluation

Alistair Plum, Felicia Körner, Anne-Marie Lutgen +8

This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for English. Although NLU tasks ar…

cs.CL2026

Languages in Whisper-Style Speech Encoders Align Both Phonetically and Semantically

Ryan Soh-Eun Shim, Domenico De Cristofaro, Chengzhi Martin Hu +2

Cross-lingual alignment in pretrained language models enables knowledge transfer across languages. Similar alignment has been reported in Whisper-style speech encoders, based on sp…

cs.CL2026

Rashid: A Cipher-Based Framework for Exploring In-Context Language Learning

Niyati Bafna, Ryan Soh-Eun Shim, Barbara Plank +2

Where there is growing interest in in-context language learning (ICLL) for unseen languages with large language models, such languages usually suffer from the lack of NLP tools, da…

cs.CL2026

SteerEval: Inference-time Interventions Strengthen Multilingual Generalization in Neural Summarization Metrics

Silvia Casola, Ryan Soh-Eun Shim, Felicia Körner +2

An increasing body of work has leveraged multilingual language models for Natural Language Generation tasks such as summarization. A major empirical bottleneck in this area is the…