most citedConcept Space Alignment in Multilingual LLMs

1 citations · 1 across the 5 of their papers we have counts for

collaborators

6 papers

cs.CL2025

Understanding Subword Compositionality of Large Language Models

Qiwei Peng, Yekun Chai, Anders Søgaard

Large language models (LLMs) take sequences of subwords as input, requiring them to effective compose subword representations into meaningful word-level representations. In this pa…

cs.CL2025

Debiasing Multilingual LLMs in Cross-lingual Latent Space

Qiwei Peng, Guimin Hu, Yekun Chai +1

Debiasing techniques such as SentDebias aim to reduce bias in large language models (LLMs). Previous studies have evaluated their cross-lingual transferability by directly applying…

cs.CL2025

SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval

Qiwei Peng, Robert Moro, Michal Gregor +7

The rapid spread of online disinformation presents a global challenge, and machine learning has been widely explored as a potential solution. However, multilingual settings and low…

cs.CL2025

Revisiting the Othello World Model Hypothesis

Yifei Yuan, Anders Søgaard

Li et al. (2023) used the Othello board game as a test case for the ability of GPT-2 to induce world models, and were followed up by Nanda et al. (2023b). We briefly discuss the or…

cs.CL20241 cited

Concept Space Alignment in Multilingual LLMs

Qiwei Peng, Anders Søgaard

Multilingual large language models (LLMs) seem to generalize somewhat across languages. We hypothesize this is a result of implicit vector space alignment. Evaluating such alignmen…

cs.CL2024

Unlocking Markets: A Multilingual Benchmark to Cross-Market Question Answering

Yifei Yuan, Yang Deng, Anders Søgaard +1

Users post numerous product-related questions on e-commerce platforms, affecting their purchase decisions. Product-related question answering (PQA) entails utilizing product-relate…