activity
20152022
most citedChallenges and Strategies in Cross-Cultural NLP

12 citations · 26 across the 8 of their papers we have counts for

collaborators

16 papers

cs.CL202212 cited

Challenges and Strategies in Cross-Cultural NLP

Daniel Hershcovich, Stella Frank, Heather Lent +11

Various efforts in the Natural Language Processing (NLP) community have been made to accommodate linguistic diversity and serve speakers of many different languages. However, it is…

cs.CL20221 cited

Finding Structural Knowledge in Multimodal-BERT

Victor Milewski, Miryam de Lhoneux, Marie-Francine Moens

In this work, we investigate the knowledge learned in the embeddings of multimodal-BERT models. More specifically, we probe their capabilities of storing the grammatical structure…

cs.CL2022

Zero-Shot Dependency Parsing with Worst-Case Aware Automated Curriculum Learning

Miryam de Lhoneux, Sheng Zhang, Anders Søgaard

Large multilingual pretrained language models such as mBERT and XLM-RoBERTa have been found to be surprisingly effective for cross-lingual transfer of syntactic parsing models (Wu…

cs.CL20211 cited

On Language Models for Creoles

Heather Lent, Emanuele Bugliarello, Miryam de Lhoneux +2

Creole languages such as Nigerian Pidgin English and Haitian Creole are under-resourced and largely ignored in the NLP literature. Creoles typically result from the fusion of a for…

cs.CL2021

Itihasa: A large-scale corpus for Sanskrit to English translation

Rahul Aralikatte, Miryam de Lhoneux, Anoop Kunchukuttan +1

This work introduces Itihasa, a large-scale translation dataset containing 93,000 pairs of Sanskrit shlokas and their English translations. The shlokas are extracted from two India…

cs.CL2020

Comparison by Conversion: Reverse-Engineering UCCA from Syntax and Lexical Semantics

Daniel Hershcovich, Nathan Schneider, Dotan Dvir +3

Building robust natural language understanding systems will require a clear characterization of whether and how various linguistic meaning representations complement each other. To…