activity
20232026
most citedWhen Is Multilinguality a Curse? Language Modeling for 250 High- and Low-Resource Languages

5 citations · 7 across the 17 of their papers we have counts for

collaborators
Showing cs.CLShow all

20 papers · 1 filter

cs.CL2026

BuzzASR: A Swarm of 100+ Monolingual Speech Recognition Models

Shivam Singh, Aditya Yadavalli, Catherine Arnett +1

We introduce BuzzASR, a collection of language-specialized fine-tuned Whisper models adapted for automatic speech recognition (ASR) in 102 languages. Large end-to-end Transformer-b…

cs.CL2026

Apples to Apples? Towards Comparable Crosslingual Language Model Evaluation

Xiulin Yang, Ethan Gotlieb Wilcox, Catherine Arnett

Crosslingual evaluation of language models that enables fair comparisons remains a fundamental challenge in multilingual NLP. Existing studies adopt a variety of downstream tasks a…

cs.CL2026

Cross-Lingual Alignment Without Joint Training: Do Monolingual Language Models Converge on Universal Representations?

Ej Zhou, Suchir Salhan, Catherine Arnett +1

Cross-lingual alignment in multilingual language models is typically attributed to joint training: shared parameters, mixed-language batches, or explicit alignment objectives. We a…

cs.CL2026

Skill Issue: Are Skills Language-Invariant in LLMs?

Bobby Cheng, Adam Gaber, Zhengyuan Liu +4

Large language models access knowledge inconsistently across languages, but to what extent do they differ in their skill sets when interacting with different languages? This work q…

cs.CL2026

Weight Tying Biases Token Embeddings Towards the Output Space

Antonio Lopardo, Avyukth Harish, Catherine Arnett +1

Weight tying, i.e. sharing parameters between input and output embedding matrices, is common practice in language model design, yet its impact on the learned embedding space remain…

cs.CL2026

How Open Must Language Models be to Enable Reliable Scientific Inference?

James A. Michaelov, Catherine Arnett, Tyler A. Chang +7

How does the extent to which a model is open or closed impact the scientific inferences that can be drawn from research that involves it? In this paper, we analyze how restrictions…